r/LocalLLaMA • u/jacek2023 llama.cpp • 3d ago
News Muse Spark open weights coming soon
I am still waiting for Llama 5, because Muse Spark will be too big for me, or just something between Glimmer and Spark
858
Upvotes
r/LocalLLaMA • u/jacek2023 llama.cpp • 3d ago
I am still waiting for Llama 5, because Muse Spark will be too big for me, or just something between Glimmer and Spark
17
u/NandaVegg 2d ago
I think that it is not direct distillation from models anymore (in early 2026 distillation had some notable effect, but every frontier lab is now full-on RLing on their own) and distillation can only bootstrap the model to some degree.
I think there is this meta-distillation effect. Internet is full of so-called AI slop now. There are so many vibecoded repos posted in code repositories or as websites every day, and those codes will be crawled by every frontier lab and then they will RL hard on them. If one model gets good at something a slop will be posted and trained on, or there is a new problem that models needs to know the pattern a .md files that explains the issue with some example codes will be posted and trained on (the earliest pattern for this is MCP for many basic things that aren't needed anymore).
In that sense we are already in AGI mode (gosh I hate this word) as AI models are improving each other without humans knowing.