mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration

Benchmark Model Rank Results
long-context-understanding-on-mmneedlemPLUG-Owl-v2#81 Image, 4*4 Stitching, Exact Accuracy: 0.3
visual-question-answering-on-mm-vetmPLUG-Owl2#100GPT-4 score: 36.3±0.1Params: 7B
visual-question-answering-vqa-on-core-mmmPLUG-Owl2#11Overall score: 20.05Deductive: 23.43Abductive: 20.6