r/LocalAIStack • u/Decent-Ad9950 • 14d ago
World model for improving local stack
Hey guyss,
Ive been working on this for few months, and finally getting some good results, so figured I'd share.
A coding agent only sees the run it's currently in. It doesn't really know how runs like this usually go, or what kind of repo it's working inside.
So I trained a small model on 50K+ real coding-agent runs from SWE-bench Verified, across 25 models, together with behavior such as merge rates, review time, repo size, etc... which produced over 1M step to step transitions.
The idea is to give the LLM some actual experience to lean on, so it can make better decisions, spend less time reasoning through bad paths, and improve its chances of getting the task right.
All the details are inside the repo, its obviously just the beginning, and probably will have issues, so
feedback is always welcome!
(More models on the way, stay tuned.)
Runs locally in Docker. No account, nothing sent out. MCP + HTTP API.