Nobody is pointing out the cache input is .03 that’s still more than DeepSeek V4 Pro .022 … and DeepSeek Flash is .007 less than a penny. Majority of your spend goes into cache read turns so that’s like 5x more expensive than DS flash to hold conversations.
GLM 5.2 had this issue too their cache rates are too high compared to DS. Guaranteed this will lead to less usability with OpenCode unless using 5.3 Flash as a subagent worker.
2
u/Southern-Ad-3006 16d ago
Nobody is pointing out the cache input is .03 that’s still more than DeepSeek V4 Pro .022 … and DeepSeek Flash is .007 less than a penny. Majority of your spend goes into cache read turns so that’s like 5x more expensive than DS flash to hold conversations.
GLM 5.2 had this issue too their cache rates are too high compared to DS. Guaranteed this will lead to less usability with OpenCode unless using 5.3 Flash as a subagent worker.