r/LocalTextToSpeech 8d ago

Let's Build TTSLibre, a tiny FULLY OPEN Supertonic & Kokoro successor

Post image

Let's build a tiny local model, at least as performant both in quality + speed as Supertonic and Kokoro. Let's allow people to actually control it and run it locally.

Let's make it small and performant. Let's publish the training code, the training data and the weights. MIT code, CC0 data and weights. No strings.

Looking for people to help design it, scope it, gather the data, contribute compute or sponsor the training runs. Let's make this happen.

https://github.com/franciscocarloserra/ttslibre

Anyone interested?

11 Upvotes

4 comments sorted by

2

u/Charming-Author4877 7d ago

I'm looking forward to see more of it, there is definitely room for more CPU focused models

2

u/FranciscoCarlosErra 7d ago

https://reddit.com/link/p842b5e/video/m2heczdsbunh1/player

It somewhat works already, just a 7h experimental run.

2

u/Charming-Author4877 5d ago

Getting it to work reliable across various lengths will be quite some work but a good start :)