Img-Diff: Contrastive Data Synthesis for Multimodal Large Language Models

Benchmark Model Rank Results
visual-question-answering-on-mm-vetMGM-7B+RPGPT-4 score: 44.1