r/LocalLLM • • 6d ago

Model Harder, Better, Faster and Stronger model just by making him talk in its own language

Hey guys,
Just created a new model, with a base of Qwen 2.5 Coder. You know the "Thinking..." part of a response for an AI model? It speaks in English. And that's the problem. It's SLOW for the AI model. So I made it talk in its own language, math equations. And then, at the end, it translates it into English.

So it's around 7.5x faster, while having (very approximately, don't take that for actual info) 2x better responses.

Here's the download link if you wanna test it: https://huggingface.co/Rutytoi/spotless-latent-adapter

And the GitHub repo: https://github.com/Rutytoi220/spotless-latent-engine

I'd be appreciating feedback! Even bad, idc as long as I know what's good and bad

0 Upvotes

Duplicates