Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models
Open paper
Benchmark
Model
Rank
Results
visual-question-answering-on-mm-vet
LLaVA-HR-X
#108
GPT-4 score: 35.5
Rank counts only results with a code link.