Video-3D LLM: Learning Position-Aware Video Representation for 3D Scene Understanding

Benchmark Model Rank Results
3d-question-answering-3d-qa-on-scanqa-test-wVideo-3D LLM#2Exact Match: 30.1CIDEr: 102.1
3d-question-answering-3d-qa-on-sqa3dVideo-3D LLM#1Exact Match: 58.6