STEP: Spatio-Temporal Progressive Learning for Video Action Detection

Benchmark Model Rank Results
action-detection-on-ucf101-24STEP#7Frame-mAP 0.5: 75Video-mAP 0.1: 83.1Video-mAP 0.2: 76.6