Contrastive Feature Masking Open-Vocabulary Vision Transformer
Open paper
Benchmark
Model
Rank
Results
open-vocabulary-object-detection-on-lvis-v1-0
CFM-ViT
–
AP novel-LVIS base training: 33.9
open-vocabulary-object-detection-on-mscoco
CFM-ViT
–
AP 0.5: 34.1
Rank counts only results with a code link.