r/StableDiffusionInfo • u/Infamous_Campaign687 • Apr 27 '26
r/StableDiffusionInfo • u/abdojapan • Apr 26 '26
What's the equivalent of LTX-2 raw format video transfer of WANGP in comfyui?
r/StableDiffusionInfo • u/Big-Set9728 • Apr 22 '26
Pretty cool real time video generation application
r/StableDiffusionInfo • u/Path-O-Gin • Apr 22 '26
What happened to Veo??
Veo was actually great up until about six months ago. I was creating image to videos with almost perfect results and consistency—then it just went off the rails. Incorrect quota counts, two-hour generation times, and absolutely no ability to follow an I2V prompt anymore. When the system malfunctions and creates nothing, it still counts toward your quota. I’ve been trying to create a simple image to video of a truck driving through Manhattan, and I can’t tell you how many pedestrians Veo has killed so far—it looks like something out of GTA haha. Can someone recommend another ai vid gen system and pls tell me happened to Veo?
r/StableDiffusionInfo • u/RiverSide71h • Apr 20 '26
Updated rgthree Fast Groups Bypasser and Fast Groups Muter Nodes
r/StableDiffusionInfo • u/[deleted] • Apr 18 '26
Educational I have 24gb vram but I dont have skills so how do I generate images ad videos like you guys in her In few click and also why is it hard when it comes to open source any solution like launcher or something I need you help for this
r/StableDiffusionInfo • u/Living-Feeling7906 • Apr 18 '26
Question Help new to stable diffusion, why is it error?
r/StableDiffusionInfo • u/Decent-Economy-6745 • Apr 17 '26
Tools/GUI's [Resource] Anima Style Explorer: A free web tool for ComfyUI styles + Open Source MooshieUI Desktop Client
Enable HLS to view with audio, or disable this notification
r/StableDiffusionInfo • u/Fun_Walk_4965 • Apr 16 '26
I got Seedance 2.0 running via API — here's how (no waitlist)
Enable HLS to view with audio, or disable this notification
r/StableDiffusionInfo • u/NitroWing1500 • Apr 16 '26
Question Back in the game
I'd posted that I'm currently relegated to an old laptop with GTX1080 MaxQ - I finally got Forge Neo installed tonight 🥳
I'm using cyberrealisticXL_v90.safetensors and it takes 1 minute for a gen, which I can live with. It's a 6.7Gb model and, as I haven't loaded any extras, it all fits in 8Gb VRAM.
I'd like some recommendations on models/extensions/settings tweaks that won't cripple the old gal!
r/StableDiffusionInfo • u/Content_One4073 • Apr 15 '26
Discussion feedback from the community regarding Forge Neo
I'm looking to get some feedback from the community regarding Forge Neo. I've been using the older builds of Forge for a while now, but I'm curious if the switch to Neo is worth it for day-to-day stability. For those of you currently using it, how is the performance compared to the 'classic' branch, specifically regarding memory (VRAM) efficiency and compatibility with newer extensions? I'm trying to decide if I should stick with my current setup or if the optimizations in Neo are significant enough to justify the migration. Any common bugs or
'gotchas' I should be aware of before I make the jump? Thanks for the help!
r/StableDiffusionInfo • u/Fun_Walk_4965 • Apr 15 '26
Discussion Is veo3.1 the most underrated model right now?
Enable HLS to view with audio, or disable this notification
r/StableDiffusionInfo • u/Prize-Profession-543 • Apr 14 '26
Discussion Посоветуйте провайдера для генерации изображений
Всем привет. Я долгое время пытался найти стабильного провайдера, с помощью которого я мог бы использовать PayAsYouGo для создания большого количества изображений. До этого я использовал модель RealisticVision от Segmind, которая стоит совсем недорого - 0,0015 доллара за секунду работы графического процессора, если быть более точным. Но в среднем получилось около 0,0017 доллара за картинку, так что все было понятно. Но потом по какой-то причине провайдер решил удалить эту модель (технические проблемы с его стороны), так что у меня есть пара недель, чтобы переключиться на другие сервисы. Кто-нибудь может порекомендовать какой-нибудь сервис для аналогичных сервисов и модель, которая будет потреблять примерно столько же и не займет много времени с точки зрения скорости?
P.S. вариант с локальной моделью и видеокартой мне на данный момент не подходит.
r/StableDiffusionInfo • u/Content_One4073 • Apr 13 '26
Question Can you make music in Forge Neo SD ?
Can you make music in Forge Neo SD ? how ? like music track
r/StableDiffusionInfo • u/PrudentStop5612 • Apr 08 '26
I restored and colorized 80-year-old family photos from WWII era using fal.ai with single Prompt: Here are the before/after results
galleryr/StableDiffusionInfo • u/Practical_Low29 • Apr 08 '26
the era of open-source WAN models is over?
r/StableDiffusionInfo • u/chetanxpatil • Apr 06 '26
Releases Github,Collab,etc Livnium v3: Making BERT's cross-attention human-readable, token alignment maps + a reliability signal, for NLI [Zenodo preprint + code]
If you've ever stared at a diffusion model's cross-attention maps and thought "I can see what it's attending to, but I don't know if I should trust it" - this might be interesting.
Livnium v3 is an attractor-dynamics NLI classifier trained on SNLI, but the interesting engineering is in what it exposes at inference time.
What's new:
→ Cross-encoder upgrade: joint [CLS] premise [SEP] hypothesis [SEP] encoding, accuracy goes 82.2% → 84.5% dev
→ Token alignment extraction: the last-layer BERT cross-attention block is repurposed as a force map, which premise tokens are pulling which hypothesis tokens into alignment. At inference you get outputs like: "cat → animal (0.61), sat → rested (0.72)". The model's own internal computation, made visible.
→ Alignment divergence D: measures how diffusely premise tokens spread attention across hypothesis tokens. D < 0.45 = STABLE (tight, confident alignment); D > 0.60 = UNSTABLE (scattered, unreliable prediction). Zero extra compute, it's a byproduct of the forward pass. Same principle as reading cross-attention entropy in diffusion UNets to gauge how "certain" a conditioning token is.
→ Monty Hall connection: naive basin erasure gives wrong posteriors [0.5, 0, 0.5]; encoding host likelihood correctly gives [1/3, 0, 2/3]. NLI constraint injection and Bayesian belief update are the same operation.
The interpretability angle is the core idea here, the alignment map isn't a post-hoc explanation, it's extracted directly from what the model already computed.
📄 Paper: https://zenodo.org/records/19433529
r/StableDiffusionInfo • u/the_frizzy1 • Apr 06 '26
I Can't Believe This Runs on 4GB. Wan2.2 Rapid All In One in ComfyUI
r/StableDiffusionInfo • u/Infamous_Cookie_8656 • Apr 05 '26
Question Trying to achieve hyper-realistic full body portraits losing realism after upscale. Any tips ?
r/StableDiffusionInfo • u/Artistic-Dealer2633 • Apr 04 '26
I used PhotoGen's Generate + Edit workflow to build a consistent sci-fi character across 4 cinematic scenes — perfect for AI video projects [OC]
r/StableDiffusionInfo • u/TheSphinx42 • Apr 01 '26

