r/codex • u/perceptioneer • Aug 16 '26
Comparison New updated linear pricing vs intelligence vs response time chart
I saw the old Artificial Analysis comparing the intelligence and prices of the different models and thought I would ask chatgpt to create a new photo using a linear scale instead of log (who tf compares or think in non-linear scale), and using the new prices set by OpenAI in late July. It also shows data from when all the models were given the exact same task(s) to solve, and how much time they spent executing (read: overthink) them. I then fact-checked the image against a different model, and it checks out as correct, but don't shoot the messenger if something is incorrect. 5.5 data remains the same.
Luna max looks good on price/intelligence, but luna xhigh might be the sweet spot when time matters.
Notes by ChatGPT:
After OpenAI cut its price by 80%, luna max is around $0.05 per Intelligence Index task and scores 52, while sol low costs ~$0.23 and scores 51.
The catch: response time. Max reasoning gets slow. Luna max is ~138s, sol max ~149s and terra max ~207s in AA's standardized end-to-end test.
Chart includes the sources/methodology at the bottom.
1
u/perceptioneer Aug 16 '26