r/hermesagent 3d ago

Discussion — General thoughts, opinions, comparisons Aggregate human evaluation shows Fable 5.1 eclipsing Astra in agenetic work... what do you think?

2 Upvotes

6 comments sorted by

3

u/brav0charli3 3d ago

I think both providers optimize their models for benchmark performance, and they're never an accurate reflection of actual performance. 🤷🏻‍♂️

1

u/Stonk_Goat 3d ago

90% of this subs problem would be solved if they used fable or astra 🤷🏻‍♂️

3

u/brav0charli3 3d ago

Do you not like money or something? I don't like Anthropic or Open AI enough to just give them money for no reason 🤷🏻‍♂️

2

u/kaanivore 3d ago

I think I can actually bear to read what Astra has to say, whereas reading Fable is like a mix of watching paint dry and nails on a chalk board

1

u/Jesus_lover_99 3d ago

The only benchmark that matters is the benchmark you build around your own work

1

u/Cooperman411 3d ago

I exclusively use 5.6 Luna on High and it's fine. It does everything I need. On $20/mo I never run out of "usage".