AutoHallusion: Automatic Generation of Hallucination Benchmarks for Vision-Language Models
Open paper
Benchmark
Model
Rank
Results
visual-question-answering-vqa-on-5
GPT-4V
#1
Overall Accuracy: 66.0
Rank counts only results with a code link.