EMNLP · 2024 · Conference paper · Top venue

Gradient Localization Improves Lifelong Pretraining of Language Models

Jared Fernandez, Yonatan Bisk, Emma Strubell

Carnegie Mellon University

Rendered by arXiv from the LaTeX source. If anything looks wrong, switch to the PDF.HTML · PDF · Open in a new tab ↗ · Close
Published in
Findings of the Association for Computational Linguistics: EMNLP 2024
Date
2024-01-01
Citations
4
arXiv
2411.04448
Cite
@inproceedings{fernandez2024gradient,
  title = {Gradient Localization Improves Lifelong Pretraining of Language Models},
  author = {Jared Fernandez and Yonatan Bisk and Emma Strubell},
  year = {2024},
  booktitle = {Findings of the Association for Computational Linguistics: EMNLP 2024},
  eprint = {2411.04448},
  archivePrefix = {arXiv},
  doi = {10.18653/v1/2024.findings-emnlp.949},
  url = {https://arxiv.org/abs/2411.04448},
}