ST-Adapter: Parameter-Efficient Image-to-Video Transfer Learning
Open paper
Benchmark
Model
Rank
Results
action-classification-on-kinetics-400
ST-Adapter (ViT-L, CLIP)
#33
Acc@1: 87.2
Acc@5: 97.6
action-recognition-in-videos-on-something
ST-Adapter (ViT-L, CLIP)
#23
Top-1 Accuracy: 72.3
Top-5 Accuracy: 93.9
GFLOPs: 8248
Rank counts only results with a code link.