OverFlow: Putting flows on top of neural transducers for better TTS
Open paper
Benchmark
Model
Rank
Results
text-to-speech-synthesis-on-ljspeech
OverFlow
#11
Audio Quality MOS: 3.37
Word Error Rate (WER): 2.30
Rank counts only results with a code link.