CVPR · 2025 · Conference paper · Top venue

CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answering

Tianyu Huai, Jie Zhou, Xingjiao Wu, Qin Chen, Qingchun Bai, Ze Zhou, Liang He

East China Normal University · Shanghai Open University · Zhejiang Zanyu Technology (China)

Rendered by arXiv from the LaTeX source. If anything looks wrong, switch to the PDF.HTML · PDF · Open in a new tab ↗ · Close
Published in
2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Date
2025-06-10
Citations
38
arXiv
2503.00413
Cite
@inproceedings{huai2025clmoe,
  title = {CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question Answering},
  author = {Tianyu Huai and Jie Zhou and Xingjiao Wu and Qin Chen and Qingchun Bai and Ze Zhou and Liang He},
  year = {2025},
  booktitle = {2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)},
  eprint = {2503.00413},
  archivePrefix = {arXiv},
  doi = {10.1109/CVPR52734.2025.01826},
  url = {https://arxiv.org/abs/2503.00413},
}