arXiv · 2025 · Preprint

Why Reinforcement Fine-Tuning Enables MLLMs Preserve Prior Knowledge Better: A Data Perspective

Zhihao Zhang, Qiaole Dong, Qi Zhang, Jun Zhao, Enyu Zhou, Zhiheng Xi, Senjie Jin, Xiaoran Fan, Yuhao Zhou, Yanwei Fu, Tao Ji, Tao Gui, Xuanjing Huang

PDF from arXiv.HTML · PDF · Open in a new tab ↗ · Close
Date
2025-06-30
Citations
14
arXiv
2506.23508
Cite
@misc{zhang2025reinforcement,
  title = {Why Reinforcement Fine-Tuning Enables MLLMs Preserve Prior Knowledge Better: A Data Perspective},
  author = {Zhihao Zhang and Qiaole Dong and Qi Zhang and Jun Zhao and Enyu Zhou and Zhiheng Xi and Senjie Jin and Xiaoran Fan and Yuhao Zhou and Yanwei Fu and Tao Ji and Tao Gui and Xuanjing Huang},
  year = {2025},
  eprint = {2506.23508},
  archivePrefix = {arXiv},
  url = {https://arxiv.org/abs/2506.23508},
}