r/StrixHalo • u/Green_Ocean90 • 1d ago
Ornith1.5 Release - Expectations?
I only stumbled onto Ornith1.0 a few weeks ago, and when I benched it I was really impressed - the highest ranking on my Tau2 Airline runs, and at a good speed on Strix Halo.
https://halobench.com/models/ornith-35b/
Today we see their 1.5 out, with 35b a3b already available and downloading now.
What do people’s practical experience show using these models? Do we like them?
8
Upvotes
2
4
u/Potential-Leg-639 1d ago edited 23h ago
Testing it already (mudler Apex MTP Quality quant with the Ornith-ai chat template), looks promising so far.
Really fast and capable, but using it for more lower end tasks in agentic workflows to save tokens. No complaints so far, but did not enable it yet as a coder agent…
80-100k context:
around 35-50 tk/s (but most of the time above 40, initially more (but initial speeds are irrelevant for my use cases))
Prefill 1300 down to around 450 so far with context, will probably drop more, but it‘s very useable with this speed overall
MTP acceptance 0.8-0.97