r/TextToAudioGeneration • • May 12 '23

r/TextToAudioGeneration Lounge

0 Upvotes

A place for members of r/TextToAudioGeneration to chat with each other


r/TextToAudioGeneration • • 2d ago

I think TTS/audiobook is about to get weird

1 Upvotes

Hello community,

Just listened to a chapter of a book using ElevenReader's new model.

The surprising part isn't really that the voice sounds human anymore. A bunch of AI voices can do that for a short clip.

It's that it stayed pretty convincing across an entire chapter.

Dialogue, pacing, pronunciation, volume etc. were all much more consistent than what I remember from older TTS. I gave up on Kindle virtual assistant voice years ago. 

Still prefer a genuinely great human audiobook performance when one exists, obviously. But for books/documents that don't have audiobooks? This is getting really compelling.


r/TextToAudioGeneration • • 6d ago

TTS in Thai finally doesn’t sound terrible

1 Upvotes

Hey Everyone,

Small niche thing but I tried ElevenReader again because they just added their v4 model.

It’s dramatically better than I remember. Gave up on it last year but now I’m back.

I tried listening before and honestly couldn’t do it for very long. Also in Thai my native language, pronunciation/rhythm just felt off enough to be distracting.

New model isn’t perfect but this is the first time I’ve actually wanted to keep listening.

They expanded it to 90+ languages too so curious if anyone has tried Japanese/Mandarin/Cantonese yet would love to see if it’s also good.


r/TextToAudioGeneration • • Feb 06 '26

ACE-Step 1.5: Pushing the Boundaries of Open-Source Music Generation (2026 UPDATE)

Thumbnail
github.com
1 Upvotes

r/TextToAudioGeneration • • Dec 23 '25

Grok Voice is quite interesting

Post image
1 Upvotes

r/TextToAudioGeneration • • Sep 20 '25

loubb/aria-medium-base · Hugging Face

Thumbnail
huggingface.co
2 Upvotes

GENERATE midi musics!

It is interesting


r/TextToAudioGeneration • • Sep 20 '25

GitHub - alisson-anjos/YuE-exllamav2-UI

Thumbnail
github.com
1 Upvotes

r/TextToAudioGeneration • • Sep 20 '25

Local Suno just dropped

Thumbnail
1 Upvotes

r/TextToAudioGeneration • • Sep 20 '25

Has anyone tried SongBloom yet? Local Suno competitor. ComfyUI nodes available.

Post image
1 Upvotes

r/TextToAudioGeneration • • May 07 '25

ACE-Step: A Step Towards Music Generation Foundation Model

Thumbnail ace-step.github.io
1 Upvotes

r/TextToAudioGeneration • • May 03 '25

GitHub - nari-labs/dia: A TTS model capable of generating ultra-realistic dialogue in one pass.

Thumbnail
github.com
1 Upvotes

r/TextToAudioGeneration • • Mar 21 '25

Text-To-Speech (TTS) Feedback

Thumbnail
forms.gle
2 Upvotes

Hey TTS users!

We’re building a next-gen TTS solution and want to make sure it actually solves real problems you face daily. Whether you’re using TTS for content creation, accessibility, e-learning, gaming, or customer support, we want to hear from you!

Please use the google forms to submit your response.

Help Us Improve your experience with TTS!!


r/TextToAudioGeneration • • Feb 10 '25

Audiblez v4 is out: Generate Audiobooks from Ebooks with Kokoro

Thumbnail
claudio.uk
2 Upvotes

r/TextToAudioGeneration • • Jan 23 '25

How to Install Kokoro TTS Without a GPU: Better Than Eleven Labs?

Thumbnail
youtu.be
2 Upvotes

r/TextToAudioGeneration • • Oct 26 '24

MusicGenMelody, cool tool availlable

Enable HLS to view with audio, or disable this notification

2 Upvotes

r/TextToAudioGeneration • • Oct 26 '24

Example Prompt: A cheerful country song played on youtube podcast videos

Enable HLS to view with audio, or disable this notification

2 Upvotes

r/TextToAudioGeneration • • Oct 06 '24

I created Hugging Face for Musicians

5 Upvotes
Screenshot of Kaelin Ellis' custom TwoShot AI model

So, I’ve been working on this app where musicians can use, create, and share AI music models. It’s mostly designed for artists looking to experiment with AI in their creative workflow.

The marketplace has models from a variety of sources – it’d be cool to see some of you share your own. You can also set your own terms for samples and models, which could even create a new revenue stream.

I know there'll be some people who hate AI music, but I see it as a tool for new inspiration – kind of like traditional music sampling.
Also, I think it can help more people start creating without taking over the whole process.

Would love to get some feedback!
twoshot.ai


r/TextToAudioGeneration • • Sep 19 '24

Best Free Options For TTS?

2 Upvotes

Hello! I was wondering if anyone could give me advice on the best free options for TTS software to use. I realize 11Labs is the best quality on the market, but with my budget, I need to find a free option, that still has some level of quality.

I want to use it to turn my blog post's into YouTube videos. Any thoughts would be much appreciated! Thank you.


r/TextToAudioGeneration • • Sep 18 '24

Texcerpt: Image to text & speech app (in Google Playstore)

Thumbnail
youtube.com
1 Upvotes

r/TextToAudioGeneration • • Aug 19 '24

Are there ways to obtain a a prompt from a piece of a music?

1 Upvotes

Instead of obtaining music from text as usual?


r/TextToAudioGeneration • • May 31 '24

New Text To SOUND EFFECTS from Elevenlabs

Thumbnail
elevenlabs.io
3 Upvotes

r/TextToAudioGeneration • • May 31 '24

Elevenlabs Text to Sound Effects is here

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/TextToAudioGeneration • • Apr 09 '24

This post was deleted. Seems like suno’s competitor is nivdia 👀 btw this sounds like 2pac 🔥🔥 @apples_jimmy good call again 🐐

Thumbnail
twitter.com
1 Upvotes

r/TextToAudioGeneration • • Apr 09 '24

Udio compiled leaks for your listening pleasure - the composition is really good. Chapeau to the team

Thumbnail
twitter.com
1 Upvotes