IET Computer Vision · 2022 · Journal article

OmDet: Large-scale vision-language multi-dataset pre-training with multimodal detection network

Tiancheng Zhao, Peng Liu, Kyusong Lee

Zhejiang International Studies University · Zhejiang University · Zhejiang University of Technology

Rendered by arXiv from the LaTeX source. If anything looks wrong, switch to the PDF.HTML · PDF · Open in a new tab ↗ · Close
Date
2024-01-24
Citations
19
arXiv
2209.05946
Cite
@article{zhao2022omdet,
  title = {OmDet: Large-scale vision-language multi-dataset pre-training with multimodal detection network},
  author = {Tiancheng Zhao and Peng Liu and Kyusong Lee},
  year = {2022},
  journal = {IET Computer Vision},
  eprint = {2209.05946},
  archivePrefix = {arXiv},
  doi = {10.1049/cvi2.12268},
  url = {https://arxiv.org/abs/2209.05946},
}