ETVA: Evaluation of Text-to-Video Alignment via Fine-Grained Question Generation and Answering

Kaisi Guan, Zhengfeng Lai, Yuchong Sun, Peng Zhang, Wei Liu, Kieran Liu, Meng Cao, Ruihua Song. ETVA: Evaluation of Text-to-Video Alignment via Fine-Grained Question Generation and Answering. In IEEE/CVF International Conference on Computer Vision, ICCV 2025, Honolulu, HI, USA, October 19-25, 2025. pages 21299-21309, IEEE, 2025. [doi]

Authors

Kaisi Guan

This author has not been identified. Look up 'Kaisi Guan' in Google

Zhengfeng Lai

This author has not been identified. Look up 'Zhengfeng Lai' in Google

Yuchong Sun

This author has not been identified. Look up 'Yuchong Sun' in Google

Peng Zhang

This author has not been identified. Look up 'Peng Zhang' in Google

Wei Liu

This author has not been identified. Look up 'Wei Liu' in Google

Kieran Liu

This author has not been identified. Look up 'Kieran Liu' in Google

Meng Cao

This author has not been identified. Look up 'Meng Cao' in Google

Ruihua Song

This author has not been identified. Look up 'Ruihua Song' in Google