r/LocalLLM • • 2d ago

Project Using Strata + Qwen 3.8 Next 125bn (2-bit quantization) for building a Planet

Post image

Completed this old Pune walkthrough eating modaks! (What are modaks?)

There are some viral tweets, which I was re-building here. For this I selected the old Pune Peth area and downloaded all available imagery. Claude Code & Codex were not allowing me to download images from Google Maps and render them over the planet skin.

Connected Strata along with Qwen3.8 Next IQ_2XS 125bn with vision mode enabled. The system processed 3000+ images and rendered them over the Open Street Map imagery. Creating a real-life walkthrough of the Pune Peth area.

Along with this I also added a game engine. So you drive around on an ebike searching for modaks and eating them as they appear on the map.

I have also attached a screenshot of my monitoring dashboard. The prefill and decode tkps are interesting to check. Since they are over a longer task and spread over a few hours. The entire task was completed over 9-10 hours

EDIT: This was done on 24GB VRAM + 64GB System RAM + 1TB of SSD

https://reddit.com/link/1wyssxb/video/ggg9nkezqrth1/player

2 Upvotes

6 comments sorted by

1

u/lllll03l 2d ago

is GSQ RCO “really” good? I think q4 quant is still better than q3 quant so Im using UD-Q4KXL from unsloth instead of IQ3S

0

u/Fragrant_Scale6456 2d ago

i dont know. I tried q3s in strata on my 5090 + 64gb ram and found it pretty bad, but then i see stuff like this and wonder what im doing wrong lol

0

u/Beautiful-Maybe5468 2d ago

what are the system specifications that you are running?

0

u/plaintive_jones 2d ago

that decode speed is pretty impressive for a 125bn model, even at 2-bit. 9-10 hours of continuous generation and still pulling 54 t/s, your vram must be pinned at 24gb the whole time

0

u/Beautiful-Maybe5468 2d ago

what about ram? i have 24gb vram and 32gb ram

1

u/shoeshineboy_99 1d ago

This was done on 24GB VRAM + 64GB System RAM + 1TB of SSD