ViLLa: Video Reasoning Segmentation with Large Language Model
Open paper
Benchmark
Model
Rank
Results
referring-expression-segmentation-on-refer-1
ViLLa
#10
J&F: 66.5
J: 64.6
F: 68.6
Rank counts only results with a code link.