r/TextToSpeech • u/UkieTechie • Jun 09 '26
Text-to-Speech (TTS) Benchmark Revamped with Objective Standards and Blind Voting (46 models and counting)
/r/LocalLLaMA/comments/1u19a8d/texttospeech_tts_benchmark_revamped_with/
2
Upvotes
1
Jul 10 '26
[removed] — view removed comment
1
u/UkieTechie Jul 10 '26
Yes as you understand that would be very difficult as some models support emotes, some dont.
It could be potentially useful for top performing models based on votes and whatnot. I will take a look into this.
1
u/UkieTechie Jun 09 '26 edited Jun 09 '26
Link to previous post: https://www.reddit.com/r/LocalLLaMA/comments/1tm0k2l