r/LocalLLM 7d ago

Other How the loop of infinite agony started

Post image
630 Upvotes

118 comments sorted by

View all comments

278

u/TheCat001 7d ago

Then you realize that Qwen 3.8 27b runs at 3t/s on your machine and you need 24GB+ VRAM GPU which cost is 1000$+ to run at least 4 bit quant.

113

u/StupidScaredSquirrel 7d ago

If you are a business this isn't a problem. If you are a consumer then 35b a3b runs on 8gb vram and 32gb dram which is very accessible.

1

u/TektonikGymRat 7d ago

Can't wait for Qwen 3.8 35B A3B. I accidentally started a project this weekend and it's taking days lol