X-Pool: Cross-Modal Language-Video Attention for Text-Video Retrieval

Satya Krishna Gorti, Noël Vouitsis, Junwei Ma, Keyvan Golestan, Maksims Volkovs, Animesh Garg, Guangwei Yu. X-Pool: Cross-Modal Language-Video Attention for Text-Video Retrieval. In IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2022, New Orleans, LA, USA, June 18-24, 2022. pages 4996-5005, IEEE, 2022. [doi]

@inproceedings{GortiVMGVGY22,
  title = {X-Pool: Cross-Modal Language-Video Attention for Text-Video Retrieval},
  author = {Satya Krishna Gorti and Noël Vouitsis and Junwei Ma and Keyvan Golestan and Maksims Volkovs and Animesh Garg and Guangwei Yu},
  year = {2022},
  doi = {10.1109/CVPR52688.2022.00495},
  url = {https://doi.org/10.1109/CVPR52688.2022.00495},
  researchr = {https://researchr.org/publication/GortiVMGVGY22},
  cites = {0},
  citedby = {0},
  pages = {4996-5005},
  booktitle = {IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2022, New Orleans, LA, USA, June 18-24, 2022},
  publisher = {IEEE},
  isbn = {978-1-6654-6946-3},
}