Vision-Language Pre-Training with Triple Contrastive Learning
Open paper
Benchmark
Model
Rank
Results
cross-modal-retrieval-on-coco-2014
TCL
#16
Text-to-image R@1: 59.0
Text-to-image R@5: 83.2
…
zero-shot-cross-modal-retrieval-on-coco-2014
TCL
#3
Image-to-text R@1: 71.4
Image-to-text R@5: 90.8
…
Rank counts only results with a code link.