Learning a Text-Video Embedding from Incomplete and Heterogeneous Data

Benchmark Model Rank Results
video-retrieval-on-lsmdcMoEE#29text-to-video R@1: 10.1text-to-video R@5: 25.6