Skip to content

Codex (gpt-5.5, Plus plan) — rate-limit cost per token jumped ~10-20x since June 16, draining the 5h budget in 2-3 prompts #28879

Description

@mihneaptu

Summary

Starting ~June 16, my ChatGPT Plus Codex budget drains in 2–3 prompts on gpt-5.5, where the same model/plan/app previously gave 20+ prompts. My Codex session logs (token_count / rate_limits events) show the limit-% consumed per token has increased roughly 10–20×, with no change on my side — prompts are actually smaller now and reasoning tokens are near zero.

Environment

  • Codex desktop app, Windows 11
  • Model: gpt-5.5, service_tier = "default"
  • Plan: plus (confirmed in rate_limits.plan_type)
  • Config unchanged across the regression window

Evidence — same plan, same model, same app

June 12 (working — 20+ prompts/day):

Prompt Input tokens Reasoning tokens Primary limit %
06:08 22,334 488 2%
06:09 52,161 1,130 3%
06:51 57,112 3,645 9% → 11%
06:54 79,961 5,125 15% → 16%

→ ~57K-token prompt with 3.6K reasoning ≈ 1% of the 5h budget

June 18 (broken — 2–3 prompts):

Prompt Input tokens Reasoning tokens Primary limit %
07:22:27 18,910 152 18%
07:22:30 20,480 0 45%
07:22:34 23,469 48 57%
07:22:43 23,631 0 68%
07:22:50 23,839 0 77%

→ ~20K-token prompt with zero reasoning ≈ 10–27% of the 5h budget

Per-token limit cost changed ~10–20×. Smaller prompts, near-zero reasoning, yet each consumes 10–20× more of the 5-hour budget than a 57K-token xhigh prompt did on June 12.

What I ruled out (verified from logs)

  • Not reasoning effort: reasoning tokens are 0–152 today vs 3,645–5,125 on the good days.
  • Not prompt/context bloat: input tokens dropped from 57K–80K to 18K–23K.
  • Not my config: model_reasoning_effort was already high/xhigh on the working days (June 12–15). xhigh gave 30–40+ prompts on June 3, 12, 13, 15.
  • Not the weekly cap: secondary window is only at 12%. The 5-hour window is the one draining.

Corroborating timing

OpenAI status page shows an active incident ongoing ~2 days (since ~June 16), which matches exactly when my usable prompt count dropped from ~9 sessions/day to ~4.

Ask

Was the Plus-tier gpt-5.5 rate-limit budget reduced or the per-token weighting changed around June 16? If intentional, what's the new budget? If not, can it be investigated?

Session log references

  • Good day: rollout-2026-06-12T09-07-36-019eba71-5c57-70d0-a8fa-aa999731e9ff.jsonl
  • Bad day: rollout-2026-06-18T10-22-14-019ed99b-ea9d-7fa1-a86b-dca4e55a8e3e.jsonl
  • Rate-limit data is in token_count events with rate_limits payload (fields: primary.used_percent, secondary.used_percent, plan_type, resets_at).

Activity

  1. added
    bugSomething isn't working
    appIssues related to the Codex desktop app
    rate-limitsIssues related to rate limits, quotas, and token usage reporting
    on Jun 18, 2026
  2. github-actions commented on Jun 18, 2026

    @github-actions
    Contributor
  3. eva763057345-lab commented on Jun 18, 2026

    @eva763057345-lab

    Same here. I'm on the Pro plan and my quota suddenly vanishes in less than 5 hours now. Even with top-ups, it doesn't survive a single intense work session. This rate-limit spike completely breaks the entire workflow. Attaching my restriction message as well (referencing image_f0485a.jpg / my previous logs). Hopefully, this gets fixed ASAP.

  4. MDGChamomile commented on Jun 18, 2026

    @MDGChamomile

    +1 here,

    seeing a sudden and unusually large drain in my Codex 5h primary limit.

    After a single simple conversational interaction, my 5h limit dropped from approximately 99% remaining to '67%'.. =(

    This seems materially higher than expected for the type of interaction performed.

    (Environment)
    Plan: Plus
    Surface: Codex Cli
    OS: Linux
    Model: gpt-5.5 high
    Approximate time: 2026-06-18 22:20~30 KST

  5. StrumykTomira commented on Jun 18, 2026

    @StrumykTomira

    Same here. Even more than that - after a few or a dozen prompts throughout the day, without any large documents (at most a dozen or so kilobytes) and without using MCP, my profile shows, for example, 22-26 MILLION (!) tokens used. And I have reasoning set to Low or, at most, Medium. This is crazy! At these prices, if I had to pay for tokens, I’d be paying for the 5.4 model (since I don’t even use 5.5) at $15 per million - that’s over $300 a day! I think they’re being calculated incorrectly. This can’t be accurate information about token usage.

  6. Aesthermortis commented on Jun 19, 2026

    @Aesthermortis

    I had to switch to Claude; now Codex seems like a more aggressive Claude 🤣. Now I feel like Claude doesn't use any resources. Codex is unusable today.

  7. andr3sdr commented on Jun 19, 2026

    @andr3sdr

    I'm experiencing the same behavior on a Pro 20x account.

    Since approximately June 18, my 5-hour limit has started draining at a rate that is completely inconsistent with historical usage. The weekly quota is decreasing proportionally as well.

    Before this change, I was never able to exhaust my 5-hour allowance, even during heavy Codex usage sessions involving large codebases and long-running tasks. Since yesterday, the same workflows are consuming the entire 5-hour budget in roughly one hour, which is much closer to the behavior I previously observed on a Plus plan.

    Nothing has changed on my side:

    • Same account and subscription tier (Pro 20x).
    • Same workflows and usage patterns.
    • No significant increase in prompt size or task complexity.
    • No changes to models or configuration.

    The timing closely matches the reports in this issue and other recent quota-related reports. From the user perspective, it appears either:

    1. The effective weighting of usage against the 5-hour budget changed around June 16–18, or
    2. The quota accounting system is currently overestimating usage.

    Given that multiple Plus and Pro users are reporting the same sudden regression over the same time window, this looks more like a platform-wide accounting or rate-limit issue than a change in individual usage patterns.

    Would be helpful to know whether any quota calculation changes were deployed around June 16–18 and whether the current behavior is expected.

  8. Truck0ff commented on Jun 19, 2026

    @Truck0ff

    Same pattern here, from a paid Pro/Codex/Hermes setup. I posted the detailed trace in #28823, but the short version is:

    2026-06-19 19:06 AEST backend endpoint:
    primary 5-hour used: 52%
    secondary weekly used: 59%
    weekly reset: 2026-06-25 08:46 AEST
    plan_type: prolite
    Spark separate bucket: 16% weekly used / 0% 5-hour used
    

    This is less than one day into the weekly window and after only small recent interaction counts. OpenAI’s public Codex pricing says Pro is “5x or 20x higher rate limits than Plus,” while Plus GPT-5.5 local messages are published as 15–80 / 5h. The practical behavior now feels materially inconsistent with that representation unless there is an undisclosed reweighting/multiplier or mis-accounting bug.

    The product also pushed an “upgrade because you are nearing limits” style message, which is not an adequate answer for an already-paid plan when the issue is unexplained quota drain.

    What users need is a per-turn/per-feature ledger: model, client/surface, bucket charged, input/cache/output/reasoning tokens, tool/MCP/image/retry/startup/compaction/background charges, and any multipliers.

    Detailed trace: #28823 (comment)

  9. avneetranjan commented on Jun 19, 2026

    @avneetranjan

    +1

  10. caioarotolo commented on Jun 19, 2026

    @caioarotolo

    Same here on a ChatGPT Pro / Pro 5x account.

    I primarily use Codex through the CLI. Over the last ~2 days, the 5-hour usage window has started draining much faster than normal, despite my workflow being basically the same as before.

  11. Alcedema commented on Jun 19, 2026

    @Alcedema

    Same issue here. Saw this start to happen when the 2x offer ended, and the remaining was not half.

  12. oscar-urbina-tech commented on Jun 19, 2026

    @oscar-urbina-tech

    This happened to me today as well. Curiously, right after I started working, I noticed my tokens were being used up very quickly relative to the five-hour usage window. I'm currently having the Pro plan.

    Does anyone know what might be going on?

  13. Vamsi-klu commented on Jun 20, 2026

    @Vamsi-klu

    +1 to this, I'm a power user in Mac. I'm seeing this since June 1 drastically increase in rate limitsvl consumption. I'm on 100 dollar plan and I see for 5 hour session I'm usually finished the limit in 1.5 hour majority of the times

  14. 188 remaining items

  15. mladenqualiteh commented on Jul 4, 2026

    @mladenqualiteh

    +1, same pattern here on Pro (100 EUR/mo): 5h meter jumped instantly from 80%+ to 0% in 1-2 messages, no warning. Also tracked in #29895 and #31125, and cross-referenced against status.openai.com incident 01KW2E6W0503W4NXJNCVAG8V6T.

  16. wackyesolution commented on Jul 8, 2026

    @wackyesolution

    i had to move on Claude, same pattern, high cost for same prompt that I send everyday. WT* are happening?

  17. raffaelkk commented on Jul 8, 2026

    @raffaelkk

    i had to move on Claude, same pattern, high cost for same prompt that I send everyday. WT* are happening?

    The squeeze?

  18. FilatovDm commented on Jul 8, 2026

    @FilatovDm

    Windows 11
    Codex Plus Plan
    GPT 5.5 Low

    I just wrote the word "Hello" and left the PC. Codex has used up 20% of the 5 hour limit and 5% of the weekly limit. For saying hello to him.
    No MCPs are enabled. There are no agents and other things. I'm using Antigravity now,
    GOODBYE Codex

  19. jyaghmour commented on Jul 8, 2026

    @jyaghmour

    Same here I am using plus plan and it is evaporating during discussion session after 10 questions

  20. lsli8888 commented on Jul 9, 2026

    @lsli8888

    Is it just me, or does Codex still feel like it's using a lot more of my quota compared to a month ago. Yes, it's better than a two weeks ago when we had a massive problem, but it still drains my quota faster nowadays.

  21. b1skit commented on Jul 9, 2026

    @b1skit

    100%.

    I have a workflow I've been using for months.

    Previously I could chat with codex for ~1 hour before being rate limited.

    Now I'm lucky to get 6 questions before using my 5hr quota, using the exact same prompts and skills.

    I still have the sessions in codex from before this started showing this huge disparity.

    Currently browsing OpenRouter for alternatives, this is ridiculous.

  22. raffaelkk commented on Jul 12, 2026

    @raffaelkk

    today the 5 hour limit is gone and since just a little bit ago limits seem to be ok for me again (im on 99% now for quite some time while yesterday some simple prompts ate a lot of pro lite....

  23. The-Cyber-Captain commented on Jul 12, 2026

    @The-Cyber-Captain

    The 5-hour limit removal is currently touted as temporary. With the ChatGPT5.6 rollout, Codex App -> ChatGPT App migration, and announcement of 5.4 retirement... it's been an absolute roller-coaster of quota resets, random performance, and your-guess-is-as-good-as-anyone regarding results weekend.

    This is blatant "move fast and break things" strategy (with a bit of Claude vs Codex PR thrown into the mix). Don't bother testing, or releasing a coherent - or even stable - product; just shite it out the door any time it suits Marketing - and have subscribers PAY to do the testing and work through the pain points. //slow-clap

  24. KaanSevinc commented on Jul 20, 2026

    @KaanSevinc

    Same Weekly Codex tokens gone in two days with a pro account.

  25. davidgilbertson commented on Jul 20, 2026

    @davidgilbertson
    Contributor

    I'm burning through my weekly limit (that used to last a week) in a day (50% of my weekly usage in ~2 hours so far today). I'm downgrading myself to Terra for now and hoping this gets fixed, otherwise I'll need to supplement with Claude.
    My 'turns per week' has dropped from ~300 previously to ~70.

  26. CanarioVidal commented on Jul 26, 2026

    @CanarioVidal

    Summary

    Starting ~June 16, my ChatGPT Plus Codex budget drains in 2–3 prompts on gpt-5.5, where the same model/plan/app previously gave 20+ prompts. My Codex session logs (token_count / rate_limits events) show the limit-% consumed per token has increased roughly 10–20×, with no change on my side — prompts are actually smaller now and reasoning tokens are near zero.
    (...)

    Same Problem!
    Win 10
    GPT Work
    4 persistent tasks

    How can I maeassure the metrics form my account like you: Prompt, Input tokens, Reasoning, tokens Primary limit %?
    Thanks!

  27. davidgilbertson commented on Jul 29, 2026

    @davidgilbertson
    Contributor
  28. FromAriel commented on Aug 27, 2026

    @FromAriel

    Cross-linking this same-model/same-plan pre/post comparison to #41220, a meta tracker for abnormal Codex usage/quota depletion and usage-accounting inconsistencies. This report is important to the cluster because the claimed ~10–20x change occurs on GPT-5.5 with smaller prompts and near-zero reasoning in the affected period, so the broader symptom cannot be dismissed as merely GPT-5.6 Sol/Ultra being expensive. The meta tracker does not presume a single root cause.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Labels

appIssues related to the Codex desktop appbugSomething isn't workingrate-limitsIssues related to rate limits, quotas, and token usage reporting

Type

No type

Projects

No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions