One Token to Seg Them All: Language Instructed Reasoning Segmentation in Videos

Benchmark Model Rank Results
referring-video-object-segmentation-on-long-rvosVideoLISA#5J&F: 33.1tIoU: 69.6vIoU: 28.2