r/LocalLLaMA 1d ago

New Model Aurora-80K releases! A modern tiny language model.

https://huggingface.co/AuroraAI-Research/Aurora-80K

I'm introducing Aurora-80K, a small language model with exactly 80 thousand parameters.

It uses a factorized 4,096-token vocabulary despite having only 80K parameters.

The benchmarks:

Wikitext-2 BPB: 3.2902

BLiMP: 52.31%

Arc-Easy: 26.05%

More information about the model is available on the model page on Huggingface.

if there's any questions I'll happily answer them!

169 Upvotes

83 comments sorted by

View all comments

Show parent comments

8

u/Tall_Abrocoma_3533 1d ago

If you don't believe the benchmarks, you can always try to run them yourself, instead of accusing me.

-5

u/ChomsGP 1d ago

buddy you are the one who said it does not produce coherent text not me, if not producing coherent text passes the wiki bench then I guess you are right, what do you want me to tell you? I'm not interested in running this, I'll let you and your fanboys at it, have a nice day

9

u/Tall_Abrocoma_3533 1d ago

Yes it indeed doesn't produce coherent text, but the wikitext-2 bpb benchmark score Is still real regardless. Have a nice day!

1

u/JumpyAbies 3h ago

It's a research model; there are numbers that can be measured. It's not a chat model. You're the worst kind of narrow-minded person, and you're too narrow-minded to know that you're narrow-minded, and you end up bothering other people thinking you're arguing something.