Region-centric Image-Language Pretraining for Open-Vocabulary Detection

Benchmark Model Rank Results
open-vocabulary-object-detection-on-lvis-v1-0DITO#2AP novel-LVIS base training: 40.4
open-vocabulary-object-detection-on-mscocoDITO#3AP 0.5: 46.1