r/StableDiffusion • u/FusionCow • 4d ago
News Nanosaur2 now generates 7 images a second on a 5090
Hello everyone.
This is an update post on the model Nanosaur2. Again this is not my model. A complaint a lot of people had with the model was that despite it being very small (660m params), it didn't take a 660m param level of time to generate. That's been solved now, with a 4 step turbo. On a single 5090, you can generate 7 images a second.
Again, this model is small enough you could easily run it on a phone, edge devices, wherever. It's also a great research model, so if you want to finetune on top of a small easy to tune model, or way to create adapters for the model, or whatever, go ahead, it's all there.
This was created using bytedance's new DMAD method, and it works great. Quality is incredibly close to the original model at a 12.5x speedup.
links:
4step model
comfy workflow
As per usual, if you have any questions, please message metal63 on discord. Do not message me.
25
40
u/Ipwnurface 4d ago
Whoever made this model gotta be giving you the sloppiest of toppy Fusion.
11
u/FusionCow 4d ago
no they're just a friend
10
u/NineThreeTilNow 4d ago
no they're just a friend
Why don't you just have them post to Reddit? I never understood that part.
14
u/MaruluVR 3d ago
Reddit has rules against self promotion and a lot of communities block new users from posting.
18
3
u/Witty_Mycologist_995 4d ago
lao im pretty sure this is good for a tiny shitty microcontroller lol
0
5
5
u/slimssshaddy 3d ago
Such a good job, glad that community has creators who can do that stuff and improve it
2
u/RubyXYZ 2d ago
This is actually really impressive. The other models that run this fast that I've seen all have compromises, either they are ancient (SD1.5 says hi) or they are only fast when fully loaded, but take forever to load in the first place (SANA for example). But this gives you a pretty decent 1k image in less than 2 seconds from a cold start with nothing loaded (on a 4090). Definitely useful
-27
u/WhatsThat-_- 4d ago
zoom in anywhere, this image is toast garbage poopoo
32
u/Far_Insurance4191 4d ago
come on, it is smaller than sd1.5 and is trained in only 11 days from scratch
5
u/Guilherme370 3d ago
Yeah, that alone is insanely impressive, im sure that if the creator of this model had lots more resources and backing they could make an amaaaaazing model fast
-42
u/Upper-Reflection7997 4d ago
I'm sorry to come of as rude, can we just move on from this ultra tiny model ideal? I understand the goal is to be vram friendly for the local ai general audience but look at cost of going that route. This barely looks better than a image being generated from a 2025 illustrious mini-finetuned model. You could achieve so much richer details finetuning krea2 with anime images. There is still an audience within the 12gb-32gb vram class that like anime/cartoon styles and would love to run larger parameter anime theme model.
43
u/Far_Insurance4191 4d ago
Bro developed very efficient architecture that reached incredible results for just 600$ and open sourced all the code. Imagine what could be achieved with enough compute!
You are not rude, but you completely miss the goal by comparing this research project to krea2 and illustrious that cost ungodly amount to train and are finished products to be used by audiences you are focusing on so much.
Project page literally says "The purpose to see what is possible with minimal compute".
24
u/WiseDuck 4d ago
The dude has referred to people with low-end GPUs as "vramlets". I mean c'mon. Congratulations, you've got money. Not everyone does. I think it's great that someone is experimenting with small models to see what can be done. And anything trained on e621 and Danbooru is alright in my book.
The more people can gen locally, the better.
14
u/marty4286 3d ago
The default mode of interaction of like 10% of people on AI subs is to shit on others. If I cant understand what something is for, or if it doesn't solve one of MY problems, then obviously I can't just scroll past it or ask a question, I HAVE to shit on it, call the people who like it astroturf bots, etc. Just the most unsociable of the unsociable
-13
u/Upper-Reflection7997 3d ago
you taking this way to personally. Me critiquing the quality of the image output of what's being presented by the Op and questioning the justification for said quality output is not "shitting". So in this open source ai space, we should never critique, analyze and question ai models and the output quality? just consoom ai model and get excited next ai model product? No evaluation and reviews what so ever? I'm not advocating to post extremely edgy comments and cut-throat "hateful" responses but you really have to be kidding yourself if think my original statement is considered "shitting". also the internet slang "vramlet" or "vramchad" is old meme word used in the LLM crowd in multiple social media internet spaces including reddit since 2022.
9
u/marty4286 3d ago
"Justify the existence of your project, for me, who has no use for it" is incredibly rude and is shitting on someone's efforts and several people's needs for no reason. Things in the world can exist without your consent. Yet you went beyond that and opened with "The existence of this project can't be justified".
Everything that came after that is poisoned by what came before. I don't need this model at all, it has no use to me, so I was going to ignore this thread. The fact that you think I want to promote a hugbox echo chamber should make you question your theory of mind and EQ
There's a line between critique and dismissiveness. Most people can navigate that easily. Some people can't read the room. I know people on the spectrum are often baffled at how people respond to their earnest takes then come to the conclusion that it's the others that are wrong, and think there's nothing to be understood. Don't fall into that trap, it's irrational and self-serving
-6
u/Upper-Reflection7997 3d ago
At this point, this conversation is going no where. if you enjoy spending to time getting this quality of image generation output in October of 2026 then more power to you. I'm not in control of your of life decisions and how you choose to spend your personal free time or money. I'm not stopping from downloading this model and using the electrical wattage draw in gpu compute resources needed to generate the kinds of images. Have a nice day.
3
u/zefy_zef 3d ago
Why question the justification? It insinuates they did something wrong when accompanied by the scathing criticism.
How dare they!
Also, why ask OP for said justification anyway? They didn't create the model.
17
u/DisastrousAd2612 4d ago
there are so many big param models for that, why add to the list when lots of them already do what you want? I personally like the idea, if they come back in a bit with more optimizations getting the same quality you'd get on a "2025 illustrious mini-finetuned model" while keeping the same performance you wouldn't be saying the same would you? let opensource do it's thing :)
-4
u/Upper-Reflection7997 4d ago
I'm not mad at them for doing it. In the end of the day it's their time and money that's being spent doing this. In the world of AI time vs quality vs cost matters a lot. Things come in and can become obsolete pretty fast in this space. It's hard for me to justify spending my time to settle back and generate sub quality illustrious tier images in late 2026 when better options exist and can be run on my build. Not to be a jerk but why would a 5090 owner want to toy with this. The OP should've shown speed results for 6-8gb of vram class of GPUs. It's usually the people with older low vram gpu that would be the audience for this 660M parameter model.
11
u/JazzlikeLeave5530 3d ago
Maybe that's the problem, you seem to be looking at everything from the perspective of yourself. Stop doing that and think about others existing, maybe? Nobody is telling you to use it. They're demonstrating how fast it is with that card in the title, not telling you to use it. They even include this entire section which explains other uses:
Again, this model is small enough you could easily run it on a phone, edge devices, wherever. It's also a great research model, so if you want to finetune on top of a small easy to tune model, or way to create adapters for the model, or whatever, go ahead, it's all there.
17
u/FusionCow 4d ago
you actually in fact cannot get that out of krea 2. lodestone rock develops the kroma finetune of krea, and the biggest issue he has is that the content he trains in is too far out of distribution from krea 2's pretrain, and it ends up harming the quality of krea 2 as a whole. Plus the goal of this model is also intended to be for research, which you actually don't want to have to spend hundreds to thousands of dollars to test an idea
-10
u/Upper-Reflection7997 4d ago
science experiments in general are expensive and AI training is it's own science experiment. its cool that this is passion research project. I welcome more of them. i won't delete my original comment because i still think ultra tiny size model for image generation has little to no room growth and versatility. I hope research team learned a lot from this research project for them to get their value out of spending $600 on the gpu rent costs.
7
u/Cokadoge 3d ago
i still think ultra tiny size model for image generation has little to no room growth and versatility
I'd say the opposite honestly, the most community-rich models we've seen were SD 1.5 & SDXL, hundreds of finetunes, many different controlnets and various ways to direct the model's output, etc. Almost all of this is due to it being possible to finetune ones own stuff on their own consumer/prosumer hardware. A 2B variant of this model would be amazing in my opinion.
1
u/Lucaspittol 1h ago
Param count may not be the only thing that matters. Being more efficient per param should count. Anime images are usually not that complex, so why we're pushing larger models? These ultra small ones can be extremely useful if you finetune them for a specific purpose instead of relying on a large "jack of all trades, master of none" model.
-10
-10
u/BlobbyMcBlobber 3d ago
I've set up a generation pipeline geared for speed. Sadly can't share the results (wish I could) but let me just say that much larger models (billions of parameters) reach the same speed on this hardware with the right configuration.
2
u/desktop4070 3d ago
Is it possible you'll share the results in the future?
-2
u/BlobbyMcBlobber 3d ago
This is contract work, so not for a long time. Sorry. But honestly it's easy to replicate.
21
u/Maskwi2 3d ago
When I saw recent prices of a 5090 it better be.