LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model

Benchmark Model Rank Results
visual-question-answering-on-mm-vetLLaVA-Phi#153GPT-4 score: 28.9