arXiv · 2024 · Preprint

MoE-CT: A Novel Approach For Large Language Models Training With Resistance To Catastrophic Forgetting

Tianhao Li, Shangjie Li, Binbin Xie, Deyi Xiong, Baosong Yang

Rendered by arXiv from the LaTeX source. If anything looks wrong, switch to the PDF.HTML · PDF · Open in a new tab ↗ · Close
Date
2024-06-25
Citations
9
arXiv
2407.00875
Cite
@misc{li2024moect,
  title = {MoE-CT: A Novel Approach For Large Language Models Training With Resistance To Catastrophic Forgetting},
  author = {Tianhao Li and Shangjie Li and Binbin Xie and Deyi Xiong and Baosong Yang},
  year = {2024},
  eprint = {2407.00875},
  archivePrefix = {arXiv},
  url = {https://arxiv.org/abs/2407.00875},
}