arXiv · 2026 · Preprint

Sparse Autoencoders enable Robust and Interpretable Fine-tuning of CLIP models

Fabian Morelli, Arnas Uselis, Ankit Sonthalia, Seong Joon Oh

PDF from arXiv.HTML · PDF · Open in a new tab ↗ · Close
Date
2026-05-15
Citations
0
arXiv
2605.15961
Cite
@misc{morelli2026sparse,
  title = {Sparse Autoencoders enable Robust and Interpretable Fine-tuning of CLIP models},
  author = {Fabian Morelli and Arnas Uselis and Ankit Sonthalia and Seong Joon Oh},
  year = {2026},
  eprint = {2605.15961},
  archivePrefix = {arXiv},
  url = {https://arxiv.org/abs/2605.15961},
}