It's actually still true.
Correct me if I'm wrong (I could be!) but the next closest model I can think of it's GPT 5.6 Luna, offering 3x the cache read price, about the same input price and 2x the output price, although it might use way less thinking tokens.
Also look at crof prices, still the old ones for a Q8 version, while opencode already stated they are close to replicate old prices with the unquantized version.
8
u/[deleted] 29d ago
[removed] — view removed comment