PreFLMR: Scaling Up Fine-Grained Late-Interaction Multi-modal Retrievers

Benchmark Model Rank Results
visual-question-answering-vqa-on-infoseekRA-VQAv2 w/ PreFLMR#1Accuracy: 30.65