TPAMI · 2023 · Journal article

Building an Open-Vocabulary Video CLIP Model With Better Architectures, Optimization and Data

Zuxuan Wu, Ze-Jia Weng, Wujian Peng, Xitong Yang, Ang Li, Larry S Davis, Yu-Gang Jiang

Rendered by arXiv from the LaTeX source. If anything looks wrong, switch to the PDF.HTML · PDF · Open in a new tab ↗ · Close
Published in
IEEE Transactions on Pattern Analysis and Machine Intelligence
Date
2023-10-08
Citations
38
arXiv
2310.05010
Cite
@article{wu2023building,
  title = {Building an Open-Vocabulary Video CLIP Model With Better Architectures, Optimization and Data},
  author = {Zuxuan Wu and Ze-Jia Weng and Wujian Peng and Xitong Yang and Ang Li and Larry S Davis and Yu-Gang Jiang},
  year = {2023},
  journal = {IEEE Transactions on Pattern Analysis and Machine Intelligence},
  eprint = {2310.05010},
  archivePrefix = {arXiv},
  doi = {10.1109/TPAMI.2024.3357503},
  url = {https://arxiv.org/abs/2310.05010},
}