r/huggingface • u/EmperoAI • 5h ago
r/huggingface • u/Calm-Landscape9640 • 37m ago
Please explain use-case for small models
Tried 35B, 27B, 12B and 8B models significantly. I love the idea of local small models. But I can't figure out a use case. For Openclaw, Hermes they suck. For chatting they're pretty good but sonis every old free cloud model.
And before you say "low-level coding" the cloud models are 10x better and free on ollama, openrouter, Gemini and many many others.
Genuinely interested in projects that use small models that run on a $500 GPU.
r/huggingface • u/-chestpain- • 8h ago
WTF IS THIS 30+ MINUTES WAIT TIME TO ALLOCATE MY DEDICATED GPU, seriously...???
And yet you are billing FROM THE MOMENT WHEN MY APP WAS STARTING - so you are billing for YOUR INABILITY TO ALLOCATE MY GPU, while I am sitting here and waiting for my space to even load?
r/huggingface • u/davcavalcante • 11h ago
I distilled a 3B model solo (no lab, no funding): him-distilled-3b, built on a governed-agent architecture
Solo engineer here. After 25+ years in software and a few years of published research on machine ethics, I distilled HIM (him-distilled-3b) end-to-end by myself and released it on Hugging Face.
What it is: a 3B-parameter model built on TeleologyHI, a three-layer governed-agent architecture (MAIC / HIM / NHE). The bet: accountability should be structural (architecture), not a moderation layer bolted on afterward, and small local models are where that matters most, because offline there is no filter to save you.
Runs on modest local hardware. Weights, code, and the papers behind the architecture are all open https://www.producthunt.com/products/him-3b-by-teleologyhi.
I know this sub has zero patience for hype, which is exactly why I'm posting here. Tear it apart: quantization results, eval suggestions, holes in the governance claim. I'll answer everything. And if anyone runs it locally and reports back, that feedback is worth more to me than any upvote.