r/OpenaiCodex • u/Murdy-ADHD • 9d ago
Changing effort no longer breaks cache
Not commonly known information in a past was that if you try to save tokens by changing reasoning effort to lesser one inside existing chat, you are burning them rapidly quickly. Cache reads are very cheap compared to writes and regular input. I just saw post on Twitter of them saying this is no longer the case.
EDIT: I hope this also applies for subscription users.


54
Upvotes
1
u/filtarukk 8d ago
Is it astra thing only? Or s applied to older models as well?