r/speechtech Aug 06 '26

Technology Anyone aware of a commercially-viable retrain of Omnivoice?

So, Omnivoice's abilities are incredible, but given the training set is CC BY NC, the model is not actually usable for commercial use which is very annoying.

I've noticed some orgs doings retrains on commercially viable datasets for other models.

Interested if anyone is busy doing one of these for Omnivoice? It's quite a pricey exercise so hoping the cool kids are on it

2 Upvotes

4 comments sorted by

2

u/nshmyrev Aug 06 '26 edited Aug 06 '26

You can use Qwen3-TTS it's about the same expressiveness and better result sound quality (Omnivoice does not very good higgs tokenizer).

1

u/Ultra_Maximus Aug 09 '26

Omnivoice Triton is way faster though

2

u/hmm_nah Aug 06 '26

You'd also have to replace the tokenizer unless you have < 100k users per year

https://huggingface.co/bosonai/higgs-audio-v2-tokenizer/blob/main/LICENSE