arXiv · 2026 · Preprint

Rosetta: Composable Native Multimodal Pretraining

Xiang-Yue Liu, Zijian Zhang, Miles Yang, Zhao Zhong, Lie-Feng Bo, Ping Tan

Rendered by arXiv from the LaTeX source. If anything looks wrong, switch to the PDF.HTML · PDF · Open in a new tab ↗ · Close
Date
2026-07-01
Citations
0
arXiv
2607.00293
Cite
@misc{liu2026rosetta,
  title = {Rosetta: Composable Native Multimodal Pretraining},
  author = {Xiang-Yue Liu and Zijian Zhang and Miles Yang and Zhao Zhong and Lie-Feng Bo and Ping Tan},
  year = {2026},
  eprint = {2607.00293},
  archivePrefix = {arXiv},
  url = {https://arxiv.org/abs/2607.00293},
}