INF-LLaVA: Dual-perspective Perception for High-Resolution Multimodal Large Language Model
Open paper
Benchmark
Model
Rank
Results
visual-question-answering-on-mm-vet
INF-LLaVA
#114
GPT-4 score: 34.5
Rank counts only results with a code link.