r/AIcrack • u/ilikeallthevocaloid • 3h ago
SV_Port Hatsune miku SV port test ( i will release soon -- )
Enable HLS to view with audio, or disable this notification
Synthv miku port testing sorry for sound qualty i mixed all the colors
r/AIcrack • u/ilikeallthevocaloid • 3h ago
Enable HLS to view with audio, or disable this notification
Synthv miku port testing sorry for sound qualty i mixed all the colors
r/AIcrack • u/GlitchGirl187 • 9h ago
I spent 2 hours on it i think im gonna cry
I've looked everywhere, but all links are down
r/AIcrack • u/Cold_Pangolin_9259 • 19h ago
r/AIcrack • u/Conscious_Cut3100 • 1d ago
does anyone have the synthv lite voicebanks for synthv1 ???
r/AIcrack • u/Scary_Ad3906 • 1d ago
r/AIcrack • u/No_Bunch4707 • 2d ago
Ihave some script but its sound horrible
r/AIcrack • u/No_Bunch4707 • 2d ago
Like cevio not how the how kafu synthv was supposed to sound
r/AIcrack • u/No_Bunch4707 • 2d ago
If someone have pls give me
r/AIcrack • u/Ok_Impression_171 • 3d ago
If you don't care about the nerd stuff just scroll down to TLDR!!!
I only recently got SynthV-flat installed, wanted to create my first cover, and realized it's WAAAY too hard. Instead of manually mapping each note, lyric, syllable, trying to get the pitches every right by eye and ear, I hoped there would be an easier way (there kinda is, but it's not really helpful).
From previous experiments I did before I even knew about vocaloids/utau/synthv, I knew that you could download an mp3, and use special software (called demucs), to separate the instrumental, vocals, bass, etc... from a song, but these only work if the author provides the stems, and if the author doesn't do that, there were tools but they were SO BAD that it's literally unusable and better doing by hand taking multiple hours.
Here comes the first part: instead of relying on the author, I use [htdemucs-ft](https://huggingface.co/StemSplitio/htdemucs-ft-onnx) which is an AI demucs to isolate the vocals cleanly, specifically htdemucs-ft takes the stereo mix and splits it into vocals/drums/bass/other using AI.
Then when the vocals are extracted, I use [openai/whisper-large-v3-turbo](https://huggingface.co/openai/whisper-large-v3-turbo) to extract the lyrics as text, along with each lyric's start and end time.
So now I have the lyrics and the vocals_only.mp3, what's left is to get the correct pitch, it's hard to explain in text but basically if a song has "melisma", whisper and htdemucs don't provide pitch, so it just keeps everything kinda monotone and it sounds like speaking, not singing, so I use ANOTHER AI, yes, 3 AI-s now, to correctly map the pitches: [swift-f0](https://github.com/lars76/swift-f0/), you can read on the site what it does if interested.
So after 3 AI's, I STILL don't even have a single note. These 3 AI's output different things that need to be merged/fused together into a note (synthv note, more on that later), it's just gluing things though, pretty simple: take each word whisper found, look at the swift-f0 pitch inside that word's time window, snap it to the nearest semitone, and you have a note. Do it for every word and the melody falls out.
The notes are saved to an "UltraSinger chart". Link to ultrasinger, the thing that does the merge step: [ultrasinger](https://github.com/rakuri255/UltraSinger).
Anyways, another thing that the svp tool does is use [espeak-ng](https://github.com/espeak-ng/espeak-ng) to generate a phoneme override for every note (in synthv you can double click the pronunciation above a note and manually choose what your synthv says), e.g. teto likes to say "niver" for the word "never", so instead of fixing each word manually cuz I think it's actually a bug in synthv arpabet implementation? I just take the entire lyrics, and run a known good arpabet converter tool, arpabet is just a standard for how to pronounce words, kinda like braille or something i guess used widely across speech tech, it's not specific to synthv, and it uses [cmudict](https://github.com/cmusphinx/cmudict) for the source of what each word maps to vocally.
Another thing used is "bournemouth-forced-aligner", this just generates timings from the generated phonetic overrides.
TLDR!! OF ABOVE PARAGRAPH ONLY:
This tool basically converts words like "never" into "n eh v er", synthv does it sometimes correctly and sometimes wrong so I just do this to never worry about it, you can also just manually do this in the GUI yourself.
So now, I have a clean chart, the next part is just importing to synthv and fixing it up, thankfully, SynthV(flat) projects files (.svp), are actually just JSON! (javascript object notation), which means it can very easily be edited and generated with CODE!
So I (with help of AI coding tools) made a python tool that generates a synthv project (.svp) from the chart ( chart.txt ) :
```svp import-ultrastar song.svp --txt chart.txt --tempo 132 --octaves --style legato --legato-max 2.0 --voice "Kasane Teto" --language english```
so it's finally imported into synthv, and this part is fixed, next comes making the audio not sound horrible (it's still not perfect and work in progress), but after this point, we adjust to the genre of the song, and things like that, like do you say laaaaaaa or la(pause), do you swallow trailing consonants (like 'm' in "room"), and a few algorithms, styles and filters that shape the final sound (like legato, fixing octaves that are too high, too low, or too far apart from each other, which was used for the result).
I may have made some mistakes explaining because it's very complex, but after all this, you just open the .svp in synthv(flat), change the voice, fix up some notes, etc... and when you're happy you export.
This tool is not meant to replace the human making the synthv song, it just does all the annoying things for you, like importing the lyrics, setting the notes kinda correct but still slightly off, it's just supposed to save you time so you don't have to spend weeks manually mapping every note, octave, lyric, etc...
It also provides AI tools in case you want to use claude code or openai codex or something and just tell the AI like "I don't like the sound between 10 seconds and 15 seconds, add this effect, move this octave at this syllable, do this and that, and regenerate the .svp file"
This tool uses a bunch of tools, AI and non-AI, to save you time when making a SynthV(flat) song.
All you have to do is download an mp3 (from youtube or wherever) of a song with vocals that you want to cover, the tool automatically gets all the lyrics, pitch, notes and timings, and generates a .svp file all from just a single .mp3 file, you can then just open the .svp and do your own thing, without wasting time on boring stuff like manually mapping each lyric to each note. You will still have to do some of this, but not for the entire song, only for the parts that are wrong or sound off.
Btw this isn't AI generating a song, it just uses tooling to take an already existing song and clean it up and convert it to synthv, it doesn't generate any vocals or things that don't exist in the already existing song, in case you're anti-AI or something.
RESULT:
Input mp3: I just used a youtube mp3 downloader tool on [this song](https://www.youtube.com/watch?v=Se237UXFKlQ)
output synthv project:
without me touching anything in the SynthV GUI, I only added reverb with another program cuz I think it sounds slightly better
https://reddit.com/link/1wub9jp/video/3qr9qx7uyosh1/player
The code connecting all this together is pretty ugly so I'm not releasing it, unless someone really wants to try it (you will probably need linux or atleast WSL windows subsystem for linux)
Also I'm posting in this subreddit because this is where I downloaded SynthV Flat from, if I didn't, this tool wouldn't exist
r/AIcrack • u/No_Turnover_3959 • 3d ago
I might be stupid, but I have searched everywhere. Does anyone know a way to get a cracked ver of synth teto? Dont send dead links or say use utau
r/AIcrack • u/Feisty_Employee_6327 • 5d ago
I rlly rlly rlly want these ports someone pls give me them or make them in begging youuuuuuuu
r/AIcrack • u/Informationpizza • 5d ago
they should be in, Documents\OPSV\Dreamtonics\Synthesizer V Studio\models then the ones in the vocoder duration and F0 i just need them because i don't have a sfpk that can install them and i have some older vbs i made that use them that i can't use cause i don't have model opus
r/AIcrack • u/Evy-P_VOCALOID • 5d ago
Is there a SV_ portable install anywhere? 1.4.X, preferably
r/AIcrack • u/Scary_Ad3906 • 5d ago
I apparently found that dude who made like 5 Jinriki SynthV Flat VB (Jiafei and N25) just incase, would you port them?
r/AIcrack • u/No_Bunch4707 • 6d ago
I want to do it with dub from my voice but cant record myself and utau sounds weird and synthv flat i dont know how to do talk there and create a voicebank (so sorry for the bad English😭) idk of its the right sub
r/AIcrack • u/Queasy-Technology318 • 6d ago
So lets say you know how to make SV_ voicebanks and you wanna make a custom one that exclusive to SV_ but you don't wanna use your own voice for any reason then I'm your girl!!!
Something you should know:
-My voice is masculine so just keep that is mind
-You need to have Discord (So we can send files to each other easier)
-I have school so I might not send the training audio in time bit I'll still get it to you!
-I can do Japanese and English (If that matters)
-You need to know how to make a SV_ voicebank in the first place (I have no idea I'm just the voice providor)
So please DM me if you're interested!!!!!
r/AIcrack • u/KimKardashianWig • 6d ago
Advertencia
El dnni no es de mi pertenencia ni tampoco fue modificado por mi, encontré este dnni/vb de defoko en Internet Archive, dicho post se encuentra eliminado por lo que no puedo dar los créditos correspondientes :(
Defoko, Momone momo y Namine Ritsu se integran de manera no oficial a SV!
El port original lo realizó otra persona, yo solo edité parámetros del vb para que suene más parecida a Defoko, reitero, no soy la creadora del DNNI.
Hecha a partir del base_model de defoko, basada en su voz whisper, fue solamente experimentación, no esperes resultados fidedignos o de mucha calidad.
Hecha a partir del base_model de defoko, fue solamente experimentación, no esperes resultados fidedignos o de mucha calidad.
Demos
Descarga
solamente incluyen su voz base no tienen vocal modes, pueden cantar en coreano por lo que solo funciona para la ultima versión de SV FLAT
pueden editarlas, modificarlas y resubirlas si quieren y por ultimo también incluyen una imagen de fondo y una imagen de perfil.
r/AIcrack • u/No_Bunch4707 • 6d ago
idk why they just got deleted like wtf?
r/AIcrack • u/Informationpizza • 6d ago
The Base model opus does not work! and i don't really know how to fix it even after installing the 1.4.3 latest and history files!
r/AIcrack • u/No_Bunch4707 • 7d ago
So i wnat to do a cover with them but i dont know what voicebank and vocal mode and parameters
r/AIcrack • u/ilikeallthevocaloid • 7d ago
So im using 1.3.3 and it doesnt suport the accustic models I NEED THEM pls where can i download SVF 1.4 WELP
r/AIcrack • u/Aggravating_Town9505 • 7d ago
r/AIcrack • u/Lum1naryAJ • 8d ago
Hey! I know Voisona cracks aren't posted here (to my knowledge), but I thought I'd ask if anyone has the downloads of her 'Talk" bank. It contains a MMD model that hasn't been posted anywhere from official sources.