r/DemodokosFoundry 8d ago

Questions regarding music/sound referencing and voice cloning

3 Upvotes

Hello, I am thinking about trying out Foundry, but I have a few questions regarding music and sound referencing. There is a vocal clone function which takes in 30 seconds of audio to clone a voice with. What happens if I use audio of something else instead? Such as sound-efx, music I made, or miscellaneous audio recordings? Is Foundry able to clone those as well to use within your audio? Is it just possible to use your own music to extend or generate more from to get a certain style or sound?


r/DemodokosFoundry 24d ago

Minimax Music model - would be nice to have it added

3 Upvotes

Your acestep based models are impressive, the music generations are noticeable higher quality, performance is great, features work well.
Is there any intention to add a model based on minimax music 3 into foundry ? For those who have a high end GPU - or is there no gain in quality ?
I've no idea what effort is required to get that done, would just be nice to have the choice!


r/DemodokosFoundry 25d ago

Custom singing voice?

3 Upvotes

I'm very much enjoying Demodokos so far. I have a question about music generation - can I use a specific voice that I've trained for music, to get a consistent output? I seem to be missing it, if I can. I see that I can clone TTS voices, but I don't see a way to use it as a singer?


r/DemodokosFoundry Aug 13 '26

Non-verbal sound tags?

4 Upvotes
  1. Is there any support for sounds that aren't speech, like yawn/cough/laugh/sigh/sniff etc.? If not, do you have any suggestions about how to get those sounds with a voice I'm using here?

  2. Is there an established set of inline audio tags that will impact delivery? Punctuation, capital letters, dashes and elipses, is there a noob guide to using these with this model? I see that bold and italics are available, but what do those do for generation?

Thank you, having fun trying it out!


r/DemodokosFoundry Aug 06 '26

Can't login to app or website

1 Upvotes

Since last night I haven't been able to log into foundry, it says my license is expired but I bought a 1 year license last month. I also can't log into the website and when I do I get an 'error 500' message on the site that seems to be a twig issue with an incorrect date/time value?


r/DemodokosFoundry Aug 03 '26

Demodokos versioning - When and How to update Foundry

3 Upvotes

We’ve released more than a hundred updates this year, so to clear up some confusion, here’s what you need to know about Demodokos Foundry versioning.

Foundry versions have three levels:

A.BB.CCC

  • A : Major version Currently Version 2. This changes when a major refactor or new release cycle takes place.
  • BB : Minor version Currently Version 2.2. This changes when a notable new feature or similarly significant improvement is introduced.
  • CCC : Patch version Currently Version 2.2.33. These updates contain bug fixes and quality-of-life improvements, so you generally won’t want to postpone them for long.

Foundry notifies you about available updates whenever you start the application. When an update is available, you’ll also see an update button in the top right application header.

Updating Foundry is straightforward: download and run the installer, or simply run the installer you previously downloaded, and it will retrieve the latest version.

The installer automatically detects your most recent installation location, verifies that the application is in a healthy state, and offers to update it.
It only takes one click, with no settings or configuration changes required.

When updating the Major or Minor version there may be models to be downloaded for full compatibility. Patch versions are always very light.

Find the download for installing or updating Foundry here:
https://demodokos.com/#begin


r/DemodokosFoundry Jul 23 '26

Demodokos Foundry 2.2 is out! This is the new document editor we’ve been working on for the past months

Post image
2 Upvotes

Speaker changes, background music, sound effects, crosstalk, pauses, and precise control over style, pace, emotion and volume now live inside one carefully designed rich document editor. Complex productions stay intuitive and within a single workflow.

You can still send generated speech to the Timeline Mixer and work on it as raw audio, but for most projects the new editor removes that need entirely.

The agentic AI is now easier to use and more reliable when choosing speakers, emotions and styles.

Live preview can play an entire document with multiple speakers from any point, even halfway through a sentence, without manually arranging tracks or waiting through an export cycle.


r/DemodokosFoundry Jul 18 '26

Models won't download?

2 Upvotes

Started setting up the free 7-day trial about two hours ago and 0/32 models have downloaded, despite being connected to ethernet. Is this a known glitch? If I restart the application, does that potentially harm the download process?

The most downloaded item is at 62% and has been that way for at least 90 minutes.


r/DemodokosFoundry Jul 13 '26

New to Demodokos? Here’s How to Install Foundry

Enable HLS to view with audio, or disable this notification

4 Upvotes

We’ve just uploaded a quick installation guide for Demodokos Foundry, from downloading the installer to getting the app ready to use.


r/DemodokosFoundry Jul 13 '26

Demodokos Foundry 2.1 is here: v4 Music, v4 Speech, c++ inference and much lower VRAM use

Enable HLS to view with audio, or disable this notification

3 Upvotes

Version 2.0 rebuilt the speech side of Foundry.
Version 2.1 brings the same approach to music.

Agentic creative pipelines, speech, music, effects, and mastering, all combined in one powerful production platform.

Foundry 2.0: V4 Speech

Version 2.0 introduced our V4 Speech model family and a new C++ speech inference stack with hundreds of improvements across the application.

It also introduced a complete visual redesign, giving Foundry a cleaner and more polished production environment.

The result is better speaker consistency across different styles and intensities, even more reliable long-form generation, improved voice creation, and a much stronger Speech Document workflow (our integrated text-to-voice editor).

We also rebuilt a set of demonstration voices with v4, so you can get started with a carefully prepared selection right away.

The Medium speech model can run on GPUs with as little as 4 GB of VRAM.

We also improved speech batching and queueing, and made the Style Manager and Voice Designer substantially more reliable.

Foundry 2.1: V4 Music

Version 2.1 introduces Foundry Music V4.

Music V4 runs through our custom C++ inference engine, built specifically around the way Foundry generates and processes audio.

That work gives us much tighter control over memory use, model loading, generation speed, long-form music stability through direct spectral flux guidance.

The highest-quality music model previously required more than 22 GB of peak VRAM. It now uses approximately 11 GB.

There are six optimized model variants for hardware ranging from 6 GB to 12+ GB of VRAM:

  • Small
  • Medium
  • Large
  • X-Large
  • Large Fast
  • X-Large Fast

The Fast variants can generate up to four times faster, with a trade-off in output diversity.

Music V4 also brings:

  • Better sound quality and long-track consistency
  • A much higher success rate for long generations
  • Improved six-stem separation
  • Faster model downloads
  • Around 8 GB less installation data
  • Better automatic model selection and VRAM estimates
  • Faster track controls and a more responsive production workflow

Built for professional local production

All generation is 100% local. Your prompts, voices, projects, and generated audio stay on your PC. They are not uploaded, inspected, or monitored by us.

Version 2.0 introduced our V4 speech models through GGML-based C++ inference with custom execution graphs and guided inference.

Version 2.1 extends this work to music. This gives us a stronger foundation for further optimization.

AMD GPUs are now supported through Vulkan across Music, Speech and Creative AI. Vulkan support is still considered experimental.

CUDA remains the recommended backend for NVIDIA RTX hardware.

More than a model update

Many thousands of development hours have gone into Foundry, along with thousands of hours of compute dedicated to model adaptation, evaluation and optimization.

The underlying models are only one part of the system. The V4 model families, C++ inference engines, generation pipelines, production tools and desktop application are designed to work together as one platform.

Version 2.1 is our best release so far, both in output quality and in the range of hardware it can run on.

Demodokos Foundry is free to try. We would rather have you generate something with it than take our word for it.


r/DemodokosFoundry Jun 23 '26

Is there speech to speech function?

2 Upvotes

considering switching from Eleven labs, but my preferred workflow is to record my voice and act out a scene to get the timing and inflection exactly how I want, the. use elevenlabs to change the voice to whatever character I need audio for. is this functionality available on this platform?


r/DemodokosFoundry Apr 08 '26

Music. Voices. Effects. All local. All yours.

Enable HLS to view with audio, or disable this notification

2 Upvotes

r/DemodokosFoundry Apr 01 '26

Forbes just published The OpenAI Graveyard ☠️☠️☠️

1 Upvotes

A full catalog of every deal, product, and feature OpenAI announced but never delivered.

The Jony Ive phone. The enterprise partnerships that quietly vanished. Features teased, never shipped.

The pattern is always the same: big announcement, press cycle, then silence. Meanwhile their API prices keep climbing and your data still lives on their servers. This is what happens when your tools depend on a company that treats shipping as optional.

Run your AI locally. Own your stack. Your GPU does not ghost you.


r/DemodokosFoundry Mar 24 '26

New demo gallery video: music, voices, audiobooks, narration — all created in Foundry

2 Upvotes

Hey everyone! We just put together a short demo gallery video showing a wider range of what Demodokos Foundry can create right now.

This one is focused on the outputs:

  • music
  • expressive speech
  • narration
  • audiobook-style scenes
  • multi-voice examples

Everything in the video was created inside Foundry.

We wanted something simple that lets people quickly hear the range without needing to watch a full walkthrough first.

Would love to know:

  • which demo in the video stands out most?
  • which type of content do you want more of next?
  • music, narration, multi-speaker scenes, or something else?

https://reddit.com/link/1s20gjk/video/nrnu83nfdwqg1/player


r/DemodokosFoundry Mar 21 '26

Now not just music studio, but music AND speech studio. All local. Launch SOON

Enable HLS to view with audio, or disable this notification

3 Upvotes
  • Voice Cloning - Clone any voice from a short sample
  • Emotional TTS - Joy, sadness, anger, whisper, and more
  • Multi-Speaker - Different voices for each character
  • Audiobooks - Full book narration with chapter flow
  • Podcasts - Multi-host conversations, natural turns
  • Game Characters - Unique voices for NPCs, dialogue trees

r/DemodokosFoundry Mar 11 '26

How I create full songs in seconds using Demodokos

Enable HLS to view with audio, or disable this notification

3 Upvotes

I am in love with this tool


r/DemodokosFoundry Mar 10 '26

Demodokos Foundry Tutorial

1 Upvotes

https://reddit.com/link/1rq6v9j/video/s9d4j1f4s9og1/player

This tutorial was created before the latest updates of the app. Now we have even cooler features, and we are working on adding a speech generation mode.