arXiv · 2026 · Preprint

Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates

Yi-Bo Li, Zijie Lin, Ailin Deng, Xuan Zhang, Yu-Fei He, Shuo Ji, Tri Cao, Bryan Hooi

Date
2026-01-26
Citations
11
arXiv
2601.18510
Cite
@misc{li2026justintime,
  title = {Just-In-Time Reinforcement Learning: Continual Learning in LLM Agents Without Gradient Updates},
  author = {Yi-Bo Li and Zijie Lin and Ailin Deng and Xuan Zhang and Yu-Fei He and Shuo Ji and Tri Cao and Bryan Hooi},
  year = {2026},
  eprint = {2601.18510},
  archivePrefix = {arXiv},
  url = {https://arxiv.org/abs/2601.18510},
}