[ACM MM2025]: MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization
Python
44
41 commits
updated Aug 13, 2025
Offical code for ACM MM2025 paper MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization (Arxiv)
π₯π₯π₯ Welcome any PR or development for MQuant!
2025.08.07: π₯π₯π₯ MQuant for Qwen2-VL has been released. Looking forward to your response!
2025.08.05: π₯π₯π₯ MQuant for Intern-VL2 and MiniCPM-V has been released.
2025.08.04: π₯π₯π₯ MQuant for Qwen-VL has been released.
2025.07.06: πππ MQuant has been accepted by ACM MM 2025. π Cheers!
Any questions or suggestions are welcome! Jiangyong Yu jiangyongyufocus@gmail.com, Sifan Zhou sifanjay@gmail.com, Dawei Yangdawei.yang@houmo.ai.
Our implementation is based on Quarot, GPTQ and VLMEvalKit. Thanks for the great open-source work!
If you think our paper or code is helpful, please consider citing our work.
@inproceedings{yu2025mquant,
title={MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization},
author={JiangYong Yu and Sifan Zhou and Dawei Yang and Shuo Wang and Shuoyu Li and Xing Hu and Chen Xu and Zukang Xu and Changyong Shu and Zhihang Yuan},
booktitle={Proceedings of the 33rd ACM international conference on multimedia (MM'25)},
year={2025}
}
MQuant is release under MIT license (see LICENSE).
Python
99.3%
[ACM MM2025]: MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization
Python
44
41 commits
updated Aug 13, 2025
Offical code for ACM MM2025 paper MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization (Arxiv)
π₯π₯π₯ Welcome any PR or development for MQuant!
2025.08.07: π₯π₯π₯ MQuant for Qwen2-VL has been released. Looking forward to your response!
2025.08.05: π₯π₯π₯ MQuant for Intern-VL2 and MiniCPM-V has been released.
2025.08.04: π₯π₯π₯ MQuant for Qwen-VL has been released.
2025.07.06: πππ MQuant has been accepted by ACM MM 2025. π Cheers!
Any questions or suggestions are welcome! Jiangyong Yu jiangyongyufocus@gmail.com, Sifan Zhou sifanjay@gmail.com, Dawei Yangdawei.yang@houmo.ai.
Our implementation is based on Quarot, GPTQ and VLMEvalKit. Thanks for the great open-source work!
If you think our paper or code is helpful, please consider citing our work.
@inproceedings{yu2025mquant,
title={MQuant: Unleashing the Inference Potential of Multimodal Large Language Models via Full Static Quantization},
author={JiangYong Yu and Sifan Zhou and Dawei Yang and Shuo Wang and Shuoyu Li and Xing Hu and Chen Xu and Zukang Xu and Changyong Shu and Zhihang Yuan},
booktitle={Proceedings of the 33rd ACM international conference on multimedia (MM'25)},
year={2025}
}
MQuant is release under MIT license (see LICENSE).
Python
99.3%