all the models — AI benchmark observatory
← Benchmarks

VATEX

VaTeX: A Large-Scale, High-Quality Multilingual Dataset for Video-and-Language Research. Contains over 41,250 videos and 825,000 captions in both English and Chinese, with over 206,000 English-Chinese parallel translation pairs. Supports multilingual video captioning and video-guided machine translation tasks.

id vatex · max 1 · 2 models reported

#ModelScore

No scores for this benchmark yet.

VATEX Leaderboard · all the models