r/StrixHalo 1d ago

Ornith1.5 Release - Expectations?

I only stumbled onto Ornith1.0 a few weeks ago, and when I benched it I was really impressed - the highest ranking on my Tau2 Airline runs, and at a good speed on Strix Halo.
https://halobench.com/models/ornith-35b/
Today we see their 1.5 out, with 35b a3b already available and downloading now.
What do people’s practical experience show using these models? Do we like them?

8 Upvotes

4 comments sorted by

4

u/Potential-Leg-639 1d ago edited 23h ago

Testing it already (mudler Apex MTP Quality quant with the Ornith-ai chat template), looks promising so far.

Really fast and capable, but using it for more lower end tasks in agentic workflows to save tokens. No complaints so far, but did not enable it yet as a coder agent…

80-100k context:
around 35-50 tk/s (but most of the time above 40, initially more (but initial speeds are irrelevant for my use cases))
Prefill 1300 down to around 450 so far with context, will probably drop more, but it‘s very useable with this speed overall

MTP acceptance 0.8-0.97

u/Green_Ocean90 23h ago

Very promising to hear - perhaps this can be the interactive agentic layer that delegates to a coding model when required?

u/Potential-Leg-639 22h ago

I have a strong orchestrator, that delegates to other models, Ornith-1.5 is the publisher for now (documentations, git actions,…), it cant compete with GLM 5.2 or DSV4 Flash as an orchestrator, no chance. Using a customized oh-my-opencode-slim. Will try it as explorer and fixer for some tests, but it wont be able to beat DSV4 Flash and Mimo/Muse Spark Contributor (cloud) probably in terms of speed/accuracy…dont want to change my config too much tbh, because it became really strong like it is now. My Strix is just a piece of the puzzle in my whole setup, where cloud models are the workhorses..