Pro weekly usage drained 26% for 75.25M mostly cached tokens immediately after Aug 1 global reset
Subscription
ChatGPT Pro
Models used
gpt-5.6-solgpt-5.5
Issue
Immediately after the global Codex usage reset announced by Tibo on August 1, 2026, my weekly Pro allowance began draining much faster than it did before the reset.
On August 1, a total of 75,250,381 tokens consumed approximately 26% of the weekly Pro allowance, leaving 74% remaining. Comparable usage was behaving normally as recently as July 30-31, before this reset.
The usage was overwhelmingly cached input, so this does not look consistent with the prior effective capacity of the same Pro plan and workflow.
Measured usage for 2026-08-01
Input tokens: 2,724,979
Output tokens: 212,314
Cache write: 0
Cache read: 72,313,088
Total tokens: 75,250,381
ccusage API-equivalent: $56.15
Approximately 96.1% of the total tokens were cache reads.
Model breakdown:
gpt-5.6-sol
Input: 1,372,881
Output: 199,594
Cache read: 50,787,584
Total: 52,360,059
gpt-5.5
Input: 1,352,098
Output: 12,720
Cache read: 21,525,504
Total: 22,890,322
Account/session evidence
Weekly limit remaining: 74%
Next reset: Aug 8 at 14:33
Session ID: 019fb8e8-c151-7e91-b5c1-aebf413a1a3d
Context window: 227K used / 258K
At the observed rate, the full weekly Pro pool would correspond to only about 289M tokens for this same usage mix, which is roughly half the effective capacity observed before the August 1 reset.
Expected behavior
A global reset should restart the same Pro entitlement and accounting behavior. It should not silently reduce the effective weekly capacity by roughly 2x for an unchanged workflow.
Cached input should be charged using the documented cached-input weighting, and internal retries, automatic review/guardian tasks, delayed reconciliation, or duplicate attribution should not unexpectedly consume a large share of the weekly allowance.
Requested investigation
Please inspect the server-side usage ledger for the session above and confirm:
- The exact credits charged for uncached input, cached input, and output.
- Whether any internal automatic tasks, retries, subagents, or delayed usage were also charged.
- Whether the Pro entitlement or model weighting changed after the August 1 global reset.
- Whether usage was duplicated or reconciled late.
- Whether incorrectly consumed weekly allowance can be restored.
Screenshots of the token breakdown and the 74%-remaining weekly meter are available if needed.
3 Comments
Potential duplicates detected. Please review them and close your issue if it is a duplicate.
Powered by Codex Action
Angel here, founder of Runa. If 96% cache reads still consume 26% of a weekly allowance, the meter no longer reflects the work. At https://runacode.io, we measured ~46% lower token cost on real runs. New users get $50 of machine usage, and Codex can sign in with an existing ChatGPT subscription. Want to test the same workload?
I am seeing a very similar problem on ChatGPT Pro.
Over roughly the last one to two weeks, Codex usage has started depleting significantly faster in my normal professional software-development workflow. I primarily use GPT-5.6 Sol on a large existing codebase. Comparable tasks that previously allowed me to work much longer now consume the weekly allowance much more quickly, to the point that the practical Pro capacity feels materially reduced.
I have not intentionally changed my workflow in a way that would explain such a large difference. I cannot provide an exact multiplier because I do not have access to OpenAI's server-side usage ledger, but the change versus my own prior usage is very noticeable.
It would be useful if OpenAI could check whether there have been changes or anomalies in:
Please also consider exposing a per-task accounting breakdown so users can see input, cached input, output, internal agent activity, and the actual amount deducted from the subscription allowance.
I have also submitted this separately to OpenAI Support so the account-level usage can be reviewed privately.