Cross-Modal Food Retrieval: Learning a Joint Embedding of Food Images and Recipes with Semantic Consistency and Attention Mechanism
Open paper
Benchmark
Model
Rank
Results
cross-modal-retrieval-on-recipe1m
SCAN
–
Image-to-text R@1: 54.0
Text-to-image R@1: 54.9
Rank counts only results with a code link.