Codex Rate Limits Discussion Thread

Hi there!
I started using Codex this week with a Plus plan to evaluate it. Today, with just three simple code-checking tasks, involving fewer than 10,000 lines, I hit the limit three times and have only 2% of my weekly quota left. Are you experiencing this kind of abnormal usage?
I normally work with Claude, and even though its consumption is high, it doesn’t come anywhere near Codex’s usage levels. With Claude the same tasks took me 6% of my weekly…

Edit:
Just checked the usage, and today I used just 1.4M Tokens. Yesterday 14.8M with no issues. Two days ago 48M hitting the limit. Always using 5.5 at mid or high, never using fast option.

i have same error

pro , 200$
but only 5 minute work

Interesting product which seems to have become a focus for what I see as the central theme of AI tools right now which is the sustainability of the business model in the face of ever-more resource hungry tools and potentially insatiable demand!
My own problem is simple and unimportant but for what it’s worth I started to trail Codex following a prompt in ChatGPT and was progressing well with a small project to evaluate its capabilities but then hit a resource cap which means I have to wait for a month under the free access model before I get to find out enough about what it can do to justify a purchase. Looking at these pages, it seems that even if I were to pay for capacity, I might continue to have the same problems, albeit on ‘the next level up’ so to speak. As it is, I’m happy to wait for a month before continuing with my trail of Codex’s capabilities but I wondered what people thought about the wider issues here? Is AI an unsustainable bubble or if not, how will it shape up?
This is ChatGPT’s summary of the situation I encountered:
“I trialled Codex on a genuine project involving cataloguing a large photograph collection. The experience was promising, but I reached the free usage limit before I had progressed far enough to determine whether Codex was actually suitable for the task. The current free allowance seems sufficient to demonstrate the interface, but perhaps not sufficient for evaluating a non-trivial workflow.”

Oh well, seems I’m not the only one…

Previously, the five-hour usage limit for Codex on the Plus plan was at least enough to go through a few rounds of prompts.

Now, after just one prompt, the five-hour limit is gone. It feels like the processing time has become longer, but the actual amount of code written has not increased. In other words, it has become slower while consuming the usage limit faster.

Before, Codex could directly modify files, and it would clearly show which files were changed and how many lines were modified.

Now, it mostly changes content through commands, making it hard to see what was actually changed.

The sandbox is already set to sandbox_mode = "workspace-write", and the project is also set as trusted.

The frequency of prompts requiring manual authorization has also increased significantly.

Hi everyone,

I’m currently on the ChatGPT Pro 5x plan and I’m experiencing severe Codex usage drain on both my 5-hour quota and weekly quota.

What makes this confusing is that my usage was normal for the first few days after I bought Pro. Then, after OpenAI said the Codex usage-limit depletion issue was fixed, my quota started draining much faster than before. This happened after the reported fix, not before.

I was planning to upgrade to the 20x Codex usage tier next month, but now I’m hesitant because I don’t know whether this is a temporary issue or not.

I’m not running unusually large tasks compared to before. The same type of Codex work that felt normal earlier now appears to consume a disproportionate amount of both 5h window and weekly usage.

Has anyone else on Pro experienced this after the incident was marked resolved?

I would especially like to know:

  • Whether others still see abnormal quota drain after the reported fix

  • Whether this could be account-specific metering or rate limiting

  • Whether Pro users were moved to a different Codex usage calculation recently

  • Whether OpenAI Support has confirmed anything for anyone affected

I wanted to ask here because the issue directly affects whether upgrading to 20x is reliable.

I’ve swapped back to 5.4 medium when using Codex, and it’s been way better for my workflow. With 5.4, I rarely hit the rate limits, but the second I use 5.5, I lock myself out in two hours and max out my weekly cap in three days.

In my opinion, 5.5 medium is only marginally better for coding tasks. By just using plan mode on 5.4 and being a little more specific with my prompts, the gap disappears. For me, not getting locked out is worth way more than a 5% difference in thinking quality.

My 5 hour limit is still depleting very quickly. It took 10 minutes to deplete while doing a fairly heavy task, but its still very unusual to deplete that fast.

I’ve been on Plus ever since, I usually hit the 5 hour limit within 2-3 hours.

OpenAI already “resolved” the issue a couple days ago, but still no improvement.

Edit:
I switched to 5.4 model and usage is somewhat normal

Title: Codex usage limits feel inconsistent after 5-hour reset — could ChatGPT Web usage affect Codex limits?

Hi everyone,

I wanted to ask if anyone else has noticed similar behavior with Codex usage limits recently.

First of all, I really like Codex and use it a lot. I have been using it for quite a while and honestly do most of my coding/workflow through it. Depending on my current workload, I switch between Plus and the higher-usage plan, so I am not against paying for higher usage when I need it.

However, recently the usage limits have felt very inconsistent and hard to understand.

What I noticed:

  • After exhausting my Codex session limit and waiting for the 5-hour reset, I sometimes do not seem to get a full session back.

  • In some cases, the next session already appears partially used, for example only around 40–50% remaining.

  • Sometimes I hit back-to-back 5-hour waits shortly after a reset.

  • This seems to happen even during periods where I am no longer using Codex directly and only use ChatGPT Web.

  • I am wondering whether ChatGPT Web code generation or Python execution might somehow count toward Codex usage, or whether the usage tracker/reset is delayed or bugged.

I tried to test this with a simple Python QR code generator in ChatGPT Web. While my Codex usage was limited, the Codex usage did not visibly change immediately. But after the 5-hour reset was over, part of my new session limit and some weekly usage already seemed to be consumed, even though I had only used ChatGPT Web during that time.

I am not saying this is definitely how it works, but something feels different from before.

My questions are:

  1. Does ChatGPT Web coding or Python execution count toward Codex session or weekly limits?

  2. Can a new 5-hour Codex session start already partially consumed because of delayed usage tracking?

  3. Were there any recent changes to Codex limit calculation for Plus or higher-usage plans?

  4. Has anyone else experienced back-to-back 5-hour waits or partially depleted sessions right after reset?

I understand that limits are necessary, especially for agentic coding tasks. But right now it feels difficult to predict what actually counts toward Codex limits, and the reset behavior is not very transparent from the user side.

Would appreciate any clarification or similar experiences from others.

I have noticed something similar recently.

In my case, I have seen GPT-4-mini usage get exhausted much faster than before. I reached the 5-hour limit in roughly one hour with a workflow that, until around two weeks ago, had never brought me close to the limit.

My main use case is code-review work: I use GPT-5 in ChatGPT for reviewing code and related development tasks, rather than only using Codex directly. Because of that, I also suspect that the token-accounting policy may have changed, or that some types of usage across ChatGPT and Codex may now be weighted or aggregated differently than before.

I cannot confirm whether ChatGPT Web usage technically counts toward Codex limits, but the practical behavior has definitely felt different recently: limits are reached earlier, resets can appear partially depleted, and it is harder to predict how much usage is actually available.

More transparency around which models, tools, and workflows contribute to the 5-hour and weekly limits would be very helpful.

Ok, i think also the github code review was wasting toons of tokens. i will see if without it everything works fine again.

Follow-up after checking my Codex session logs

I checked the exported JSONL logs for the 12 Codex Desktop sessions I ran on July 4–5.

All 12 sessions explicitly report:

  • Model: gpt-5.4-mini
  • Reasoning effort: high
  • No visible gpt-5.5 calls
  • No visible subagents, spawned agents, or nested tasks

However, the account usage dashboard for July 5 shows:

  • gpt-5.4-mini: 19 turns
  • gpt-5.5: 31 turns
  • gpt-5.4: 2 turns

I am attaching the screenshot because this creates a real attribution problem.

The Codex Desktop logs themselves do not explain the GPT-5.5 usage. I do use GPT-5.5 in ChatGPT for manual code review, planning, and product work, so one possibility is that the dashboard aggregates ChatGPT Web, Codex Desktop, Code Review, or other surfaces into the same usage picture. Another possibility is that some workflow uses a different model behind the scenes without exposing that clearly in the local Codex session logs.

I cannot prove from this that GPT-5.5 ChatGPT usage directly consumes the same Codex quota. But I can confirm that, on the same day, my local Codex task logs show only GPT-5.4-mini while the usage dashboard records substantially more GPT-5.5 turns.

That makes it very difficult to understand why a 5-hour limit is exhausted, especially when the visible Codex work appears to be using a smaller model.

It would be very helpful to have:

  • a breakdown by surface: ChatGPT Web, Codex Desktop, Code Review, Cloud Tasks, etc.;
  • confirmation of which surfaces share the 5-hour and weekly pools;
  • any model or task weighting applied to limits;
  • visibility into whether subagents or background model calls are counted;
  • an explanation of what activity contributed to a specific limit reset or exhaustion.

Without that, users cannot reliably tell whether they are consuming Codex capacity through coding tasks, ChatGPT Web reviews, automated code review, or some combination of all of them.

$200 Pro user issue: Over the past two weeks, my weekly quota has been consumed at an extremely high rate—almost 40% or more per day, depleting it completely within just two days. This was never possible before. Additionally, I usually have over 50% of my five-hour quota remaining each time, yet my weekly quota is being used up so quickly. I’m unsure why this is happening and would appreciate a response.

Hi and welcome to the developer community.

Do you use goal, subagents, speed?

Hello, I’ve also been running into very strange rate limits over the last few days. Notably, yesterday I used 93% of my 5H usage across 7 messages using ~500k tokens. I’m on a Pro 5x plan, and before this week I’ve never run into issues like this. I’m not sure where to go to find out why this is happening, but would really appreciate some sort of explanation

I am an active ChatGPT Pro subscriber paying $200/month. Before the latest updates, I could run heavy development workflows and agentic tasks (like OpenAI Operator) within the flat-rate membership.

Now, the exact same tasks burn through limits aggressively and dump me into a metered billing loop. This month alone, I have had to pay an extra $600 in credit top-ups just to keep my daily work from freezing.

The math makes no sense. I could literally buy THREE separate $200 Pro accounts for $600 and get three times the baseline token volume, rather than paying for individual metered credits. However, managing three separate accounts completely breaks codebase context and workspace memory.

Forcing premium $200 users into aggressive micro-transactions completely destroys the value proposition of the Pro tier. Is anyone else experiencing massive price inflation on workflows that used to be covered? OpenAI needs a true flat-rate developer tier that doesn’t restrict agentic workflows with a text meter.

I still cannot understand what is happening with Codex. The limits are being burned at an absurd rate.

Before the GPT-5.6 series, a $20 subscription let me work with GPT-5.5 for weeks without constantly thinking about usage. I did not even know what it felt like to hit the weekly limit. I simply never got there.

I also have measurements from a particularly busy week before GPT-5.6. I still never hit the 5-hour limit, and the weekly limit was not even close. There was roughly 30-40% left.

I went back through my chat history and found screenshots I took on June 11, 2026:

  • 10:20: before the weekly reset, 89% of the 5-hour limit and 47% of the weekly limit remained.

  • 10:25: the weekly limit reset.

  • 10:54: after the reset, 37% of the 5-hour limit and 94% of the weekly limit remained.

The weekly counter reset to 100%, while the 5-hour counter did not. Since the weekly limit was about to reset anyway, I decided to use GPT-5.5 Extra High to implement a large feature. It created 48 files, wrote about 1,600 lines of code, and worked for 40 minutes. The whole task used 8.3% of the weekly limit and 52% of the 5-hour limit.

And now? Two GPT-5.5 Medium prompts where I described exactly what needed to change and where. Six minutes of model work in total, including build time, and 180 lines of code. Result: 4% of the weekly limit gone.

I will not even get into Sol Low and Medium. This morning I gave it a task to check an already completed solution against the requirements. It worked for 23 minutes, added 130 lines, deleted one file, and used one Terra Low subagent as a code formatter. Around 14% of the weekly limit disappeared. This is not usable.

Over the last three months with Codex, I have never used an effort level above High on any model except Luna, where I sometimes used xHigh or Max. Most of the time I use Medium for normal work and Low for simple tasks. I have probably used High only 3-4 times in total. I always clear context or use /compact when old context is no longer relevant. I genuinely do not know what else I am supposed to do to preserve my limits.

I understand that usage is not measured by lines of code or wall-clock time alone. Context size, reasoning, tool calls, subagents, and output all matter. But that is also the problem: the practical cost of a task has become extremely difficult to predict.

I see people with $100 and $200 subscriptions saying they burn through their limits in 2-3 days. Tibo’s resets were the only thing saving them. Is that really considered normal now?

People are already making X skills to reduce usage. sol-advisor, for example, delegates work to Luna. Yes, that makes limits drain more slowly, roughly at the level they did before GPT-5.6. But what is the actual answer here? Are we supposed to use a smaller model just to get back to the old usage level?

For me, that feels either like a very transparent “pay more” message or like a seriously bad release. There were also plenty of API and Codex issues at launch.

Luna is not a bad model in general, but it does not work well for the coding tasks I use it for. It is only usable on Max or xHigh reasoning, which makes it incredibly slow. Its knowledge is not enough for me to rely on it the way I used to rely on GPT-5.5. It often writes overly complicated, poor code and agrees with almost any questionable approach.

In one case, Luna Max, launched through sol-advisor with Sol High, wrote around 300 lines of code across 12 files, plus 7 additional empty files. The entire stack still consumed about 9% of my weekly limit.

I also tried using Opus 5 High for an entire day. I did not hit the 5-hour limit even once. I know a group of five people who share a single $100 subscription to Claude. They have never hit the weekly limit and only hit the 5-hour limit a couple of times. At the end of each weekly reset, they still have more than 40% left.

Based on my own usage, I strongly suspect that Sol Low would use their weekly limit much faster.

All the posts saying, “I am a Claude user, convince me to switch to Codex,” followed by replies about large limits and Tibo resets feel almost funny now. The resets will not last forever, and the large limits seem to be gone already.

Please, if you have the same problem, like or reply to this post. And if you are not seeing it, please share your own observations.

I really need to know whether I am losing my mind here :smiling_face_with_tear:

I don’t understand what’s happening anymore. The 5‑hour usage limit “reset” is gone now it only weekly and th daily limit are done keeps pushing the remaining time 2–3 days further every single day. Today I added $80, and four hours later my balance dropped to $12 permanently. The model stalls for 20+ minutes just to tell me to log in, and I can’t even access or view my edits. At this point, I’ve spent more this month than it would cost to activate two additional Pro Max subscriptions — and the system still isn’t functioning properly.

I purchased a second Pro $200 plan and managed to exhaust my entire week of usage within a single day and only 38.2M tokens. And yeah, I get it that if I used token billing that’d be $190… But this is absurdly low usage and a minuscule fraction compared to what I am used to from several years using OpenAI models.

I’ve noticed a major difference between the Codex usage included with my subscription and the additional credits purchased separately.

The included subscription allowance appears to go considerably further. By comparison, the £20 bundle of 500 credits can disappear extremely quickly, sometimes during fairly ordinary website-editing tasks.

The bigger problem is that I often cannot see my usage at all. The usage information is either unavailable, missing or not detailed enough to show what each task has cost. That makes it impossible to understand why credits have fallen so quickly or whether a failed, incomplete or repeated task has still been charged in full.

Customers should be able to see:

  • the credits used by each task;
  • the model and mode responsible for the charge;
  • whether retries and failed actions consumed credits;
  • the remaining subscription allowance;
  • the remaining purchased-credit balance;
  • a clear history of when credits were deducted.

Customer service has also been poor in my experience. I have contacted support about credit usage and incomplete tasks, but messages are frequently ignored. When I do receive a reply, it often fails to address the specific questions raised.

It is not reasonable to sell additional credits without providing a dependable usage record or responsive support when those credits disappear unexpectedly.

Has anyone else found that their subscription allowance lasts much longer than purchased credits, while the usage information is unavailable most of the time?