Learning Language-Visual Embedding for Movie Understanding with Natural-Language
Open paper
Benchmark
Model
Rank
Results
video-retrieval-on-msr-vtt
C+LSTM+SA+FC7
–
text-to-video R@1: 4.2
text-to-video R@10: 19.9
…
Rank counts only results with a code link.