Audio-Visual Activity Guided Cross-Modal Identity Association for Active Speaker Detection
Open paper
Benchmark
Model
Rank
Results
audio-visual-active-speaker-detection-on-ava-activespeaker
GSCMIA
#9
validation mean average precision: 92.86%
Rank counts only results with a code link.