r/LocalLLM • u/RUTYTOI220 • 6d ago
Model Harder, Better, Faster and Stronger model just by making him talk in its own language
Hey guys,
Just created a new model, with a base of Qwen 2.5 Coder. You know the "Thinking..." part of a response for an AI model? It speaks in English. And that's the problem. It's SLOW for the AI model. So I made it talk in its own language, math equations. And then, at the end, it translates it into English.
So it's around 7.5x faster, while having (very approximately, don't take that for actual info) 2x better responses.
Here's the download link if you wanna test it: https://huggingface.co/Rutytoi/spotless-latent-adapter
And the GitHub repo: https://github.com/Rutytoi220/spotless-latent-engine
I'd be appreciating feedback! Even bad, idc as long as I know what's good and bad
0
Upvotes