ActBERT: Learning Global-Local Video-Text Representations
Open paper
Benchmark
Model
Rank
Results
action-segmentation-on-coin
ActBERT
#7
Frame accuracy: 57.0
Rank counts only results with a code link.