Learning Cross-modal Context Graph for Visual Grounding

Benchmark Model Rank Results
phrase-grounding-on-flickr30k-entities-testLCMCG#6R@1: 76.74