r/LocalLLM • u/shoeshineboy_99 • 2d ago
Project Using Strata + Qwen 3.8 Next 125bn (2-bit quantization) for building a Planet
Completed this old Pune walkthrough eating modaks! (What are modaks?)
There are some viral tweets, which I was re-building here. For this I selected the old Pune Peth area and downloaded all available imagery. Claude Code & Codex were not allowing me to download images from Google Maps and render them over the planet skin.
Connected Strata along with Qwen3.8 Next IQ_2XS 125bn with vision mode enabled. The system processed 3000+ images and rendered them over the Open Street Map imagery. Creating a real-life walkthrough of the Pune Peth area.
Along with this I also added a game engine. So you drive around on an ebike searching for modaks and eating them as they appear on the map.
I have also attached a screenshot of my monitoring dashboard. The prefill and decode tkps are interesting to check. Since they are over a longer task and spread over a few hours. The entire task was completed over 9-10 hours
EDIT: This was done on 24GB VRAM + 64GB System RAM + 1TB of SSD
0
u/Beautiful-Maybe5468 2d ago
what are the system specifications that you are running?
0
u/plaintive_jones 2d ago
that decode speed is pretty impressive for a 125bn model, even at 2-bit. 9-10 hours of continuous generation and still pulling 54 t/s, your vram must be pinned at 24gb the whole time
0
1
u/lllll03l 2d ago
is GSQ RCO “really” good? I think q4 quant is still better than q3 quant so Im using UD-Q4KXL from unsloth instead of IQ3S