Justin McKelvey

Justin McKelvey

Fractional CTO · 15 years, 50+ products shipped

• Vibe Code Rescue • 10 min read •

Codex Usage Limits Explained: The 5-Hour Window, the Weekly Cap, and What Actually Resets (2026)

Quick Answer

Codex usage limits are a rolling five-hour allowance per model, with a weekly cap that may sit on top, and OpenAI publishes ranges rather than numbers. As of September 2026 its own page estimates a Plus ($20/month) user gets roughly 10–100 GPT-5.6 Sol messages per five-hour window (25–200 on Terra, 250–2,000 on Luna), Pro 5x ($100) about five times that, and Pro 20x ($200) about twenty times. Cloud chats burn more than local ones, fast mode and images burn faster still, and when you hit the wall mid-task the agent finishes the turn. The number to remember: the meter counts tokens, not messages.

Verified September 2026 against OpenAI's Codex pricing page · Author: Justin McKelvey, fractional CTO & founder of Vibe Code Rescue, 15 years in software, 50+ products shipped

TL;DR: Codex Usage Limits Are a Window, Not a Count

The question "what are the Codex usage limits" has no single-number answer, and that is by design. OpenAI's pricing page, read today, describes an allowance that refills on a rolling five-hour window, varies by which model you run, may also be capped weekly, and is spent by tokens rather than by messages. Every ChatGPT plan gets a multiple of the same base allowance: Plus is the base, Pro 5x is five times it, Pro 20x is twenty times, a standard Business seat matches Plus, and the $100 Business seat matches Pro 5x. Nothing here is a fixed cap, which is why two people on the same plan report wildly different experiences. One runs Luna for quick edits and never sees the limit; the other keeps a 40-file repo in context on Sol and hits it before lunch.

I run Codex, Claude Code, and Cursor side by side on real client work, and the honest version of this post is that the limits stopped being the interesting question once I understood the meter. Here is the meter.

The Codex limits by plan, from OpenAI's own page (September 2026)

These are OpenAI's published estimates of local messages per five-hour period, re-read on September 28, 2026, and the page says out loud that they are estimates, not fixed message limits. Cloud chats run GPT-5.6 Sol by default and may consume more of the allowance than a local message does. A standard Business seat gets exactly the Plus column.

Model Plus ($20) and standard Business Pro 5x ($100) Pro 20x ($200)
GPT-6 Astra 5–45 25–225 100–900
GPT-6 Sol 15–150 70–700 300–3,000
GPT-6 Luna 350–3,000 1,750–14,000 7,000–56,000
GPT-5.6 Sol 10–100 50–500 200–2,000
GPT-5.6 Terra 25–200 125–1,000 500–4,000
GPT-5.6 Luna 250–2,000 1,250–10,000 5,000–40,000

The page also still carries rows for GPT-5.5, GPT-5.4 and GPT-5.4 mini, and says GPT-5.5 retires from ChatGPT, ChatGPT Work and Codex on every plan on October 14, 2026, so I've left the older models out of the table.

Three lines of fine print matter more than the table. A standard Business seat ($20 per user per month billed annually, $25 monthly, two-user minimum) gets the Plus estimates; "Business ($100) uses the Pro 5x estimates." "Weekly limits may also apply" is the page's exact phrasing, with no weekly number published. And "ChatGPT Work and Codex share usage", meaning agentic work you run inside ChatGPT draws down the same allowance as your CLI sessions. The Go tier ($8/month) exists for lightweight coding and is not in the estimate table; the API key route has no plan windows at all and bills at usage rates instead.

One Astra message, in other models' messages

OpenAI never says how the models compare, but its own ranges do. Divide each model's range by GPT-6 Astra's on the same plan and you get an exchange rate for the window. The ratios hold on every plan, because Pro 5x and Pro 20x are exact multiples of Plus.

Model Plus range Messages worth one Astra message (small tasks / large tasks)
GPT-6 Astra 5–45 1 / 1
GPT-5.6 Sol 10–100 2.2 / 2
GPT-6 Sol 15–150 3.3 / 3
GPT-5.6 Terra 25–200 4.4 / 5
GPT-5.6 Luna 250–2,000 44 / 50
GPT-6 Luna 350–3,000 67 / 70

A worked morning, to make that concrete. Say you send ten Codex turns on Plus. On Astra, ten turns are 22% of a window if your tasks are small enough to sit at the top of OpenAI's range (45 per window), and two full windows if they're big enough to sit at the bottom (5 per window), which is ten hours of waiting. The same ten turns on GPT-6 Sol are 7% to 67% of one window. The assumption doing the work: a task that sits at the top of one model's range sits at the top of the others' too. OpenAI doesn't state that, but it is the only way to read ranges published side by side.

The practical rule it gives you: each Astra turn costs about three GPT-6 Sol turns, so keep it for the turn that needs it and let GPT-6 Sol or Luna do the rest.

Codex on Free and Go, and the "monthly limit" question

Free ($0) and Go ($8 a month) are both listed as "GPT-6 Luna at Standard speed in the desktop app, subject to rollout." Neither gets a column in the estimate table, so OpenAI publishes no per-window number for either; the web, CLI, IDE and iOS surfaces and the cloud integrations are listed on the Plus card, not on these two. Image generation isn't available on Free at all.

There is no monthly Codex limit on a ChatGPT plan. The two clocks are the five-hour window and the weekly cap. The only monthly-shaped meters are on the business side: Enterprise and Edu workspaces with flexible pricing have "no fixed rate limits" and spend credits instead, and without flexible pricing they get "the same per-seat usage limits as Plus."

The 5-hour window: what resets and when

The window is rolling, not calendar-based. Usage you spent at 9am frees up at 2pm regardless of what you did in between, which is why a heavy morning and a heavy afternoon feel like two separate budgets and a heavy single stretch feels like a wall. The ChatGPT usage dashboard shows current limits and reset times; inside a Codex CLI session, /status shows the same thing without leaving the terminal. If you have never typed that command, type it before you decide you need a bigger plan.

The Codex weekly limit

OpenAI states that a weekly cap may apply on top of the five-hour allowance and does not publish it, which is the single most-searched frustration about this product and the reason "codex weekly limit" is its own query. Read the two meters together and the order of events is predictable: on an ordinary week the five-hour window is the one you meet, and the weekly cap only shows itself after several days of long agentic runs in a row, at which point the five-hour reset stops helping. If you are hitting it, you are a heavy user by any definition, and the plan question below is a real one for you rather than a hypothetical.

Codex rate limits vs usage limits

The two phrases get used for the same thing, and the page itself mixes them. The rolling window and the weekly cap are what a ChatGPT-plan user means by either. The other meaning, rate limits in the API sense (requests and tokens per minute by API tier), only applies if you authenticate Codex with an API key instead of a ChatGPT login. That route drops the plan windows entirely, bills at API pricing, and gives up the cloud features (GitHub code review, Slack) that live on the ChatGPT side. It is the right choice for CI and shared environments and the wrong one for a person who wants a predictable monthly bill.

What burns the allowance faster than you think

  • Context. A session holding a large codebase costs more per turn than a fresh one; the page lists context, reasoning, tool use, retrieval, and caching as the variables, and says prompt length alone is not a reliable estimate.
  • Cloud chats. They run Sol by default and may use more of the allowance than local messages.
  • Fast mode. Speed configurations consume credits at a higher rate and therefore use included limits faster.
  • Images. Image generation draws on the same limits roughly three to five times faster on average.
  • Voice. Voice in Desktop now spends your existing Codex budget "at $0.05 per minute," and the model doing the task is billed on top at its normal rate. On credit-billed workspaces it's 1.25 credits a minute. (Earlier this month the page gave Voice its own allowance and listed a separate Codex-Spark limit; neither appears on the September 28 read.)

What happens when you hit the limit mid-task

The page is explicit that work in progress gets to finish: if you reach the limit during an active turn, the agent continues that turn subject to fair use. So a long migration does not die halfway; it stops accepting the next instruction. From there you have four moves, cheapest first: wait for the window, switch to a smaller model (Terra or Luna instead of Sol) so what remains lasts longer, buy additional credits (Plus and Pro can do this without changing plan), or run extra local chats against an API key at usage rates. Upgrading is the fourth option, and it is the one people reach for first.

Should you upgrade to Pro for the limits?

Decide it from the meter, not from a bad afternoon. OpenAI's page itself suggests checking the usage dashboard "every week or two" to learn your pace. Run /status at the end of three normal working days and note how often you were actually blocked versus annoyed. If Plus blocks you most days, Pro 5x at $100 is the honest tier, and it is worth knowing that OpenAI now sells that middle rung; the "$20 then $200 cliff" I described in my Codex pricing breakdown has a $100 step in it as of September 2026. If you are blocked once a week, buy credits and keep the $20 plan. If you are blocked because every session starts by re-explaining the whole codebase, fix the AGENTS.md and the task scoping first; how to use Codex covers both, and it is free. The same three-step order (threads, scoping, then money) is what I wrote for the Claude side in Claude usage limits. If you're weighing a switch rather than an upgrade, Claude's, Gemini's and Copilot's ceilings sit next to OpenAI's in the AI usage limits by plan table.

Codex rate limits: what resets and when

"Codex rate limits" is the phrase people type. The thing they are hitting is the plan allowance above, so here is the reset behavior in one place, from OpenAI's Codex pricing page and its Codex help article, both read on September 15, 2026. There are two clocks. The five-hour window refills on its own schedule, and the usage dashboard shows the exact reset time. The weekly limit resets on a fixed date; OpenAI's help center confirms both exist by describing how a banked reset "refreshes your 5-hour and weekly Codex usage windows and changes your weekly reset date." OpenAI publishes neither the weekly number nor a reset day.

What hitting one looks like: OpenAI publishes no exact error string. In the app you get a limit banner that names which allowance is exhausted and lists your options: add credits, apply an available reset, upgrade, or wait for the displayed reset time. In the CLI, /status shows token usage and /usage shows daily, weekly, or cumulative activity and lets you redeem an earned reset. Support will not reset a limit for you; the help center says so directly.

Two things that are not this limit. API rate limits (requests and tokens per minute by tier) belong to an API key and only apply if you signed in with one. And ChatGPT's own caps on images, uploads, and voice have separate banners and, in OpenAI's words, "do not apply to Codex."

Codex limits when you're paying for the team

If you own the company rather than write the code, here is the version I give on calls. A developer hitting Codex limits every day is either doing genuinely heavy work, in which case $100 a month is the cheapest engineer-hour you will ever buy, or is running the same kind of task by hand every morning, in which case the limit is a symptom and the fix is a workflow, not a plan tier. The AI Readiness Assessment ($2,500, credited in full against a build within 90 days) is how I tell those two apart in two weeks. And whatever tier you land on: before real users touch what Codex shipped, run it against the free 20-point vibe coding security checklist. Unreviewed agent output under a generous allowance is exactly how apps end up in my rescue queue.

Related: Codex pricing · is Codex free? · how to use Codex · Codex review · Cursor vs Claude Code vs Codex · Claude usage limits.

The Pro 20x column you can't buy right now

Noted September 20, 2026: since September 10, OpenAI has not taken new sign-ups or upgrades to Pro 20x ($200/month); GPT-6 Astra demand after its September 3 launch outran capacity. For Codex that means the right-hand column of the table above, 100–900 Astra or 200–2,000 GPT-5.6 Sol messages a window, belongs only to people who already had it. Those subscriptions keep renewing, but anyone who cancels or downgrades can't come back until the pause lifts, and OpenAI hasn't said when that is. The biggest Codex allowance you can actually buy today is Pro 5x at $100, 25–225 Astra or 70–700 GPT-6 Sol a window, and past that the options are credits or a second seat.

Sources

  • OpenAI, Codex pricing (plans, per-model estimates, weekly-limit and fair-use wording, Voice, credits): read September 8, 15 and 28, 2026
  • OpenAI Help, About ChatGPT Pro plans (the Pro 20x pause): read September 20, 2026
  • OpenAI Help Center, Codex usage and resets article (banked resets, no Support resets, /status and /usage): read September 15, 2026

Next step Get the free cost sheet →

Free Resource Justin McKelvey

What your AI stack actually costs

Prices changed 3x this year. The always-current cost sheet: sticker price vs what heavy use actually costs for Cursor, Claude, Replit, Lovable, Bolt & more.

Frequently Asked Questions

What are the Codex usage limits?
As of September 2026 OpenAI describes Codex limits as an allowance per rolling five-hour window, not a fixed message count. Its own estimates for local messages per five-hour window on Plus ($20/mo) are roughly 10–100 with GPT-5.6 Sol, 25–200 with Terra, 250–2,000 with Luna, and 5–45 with GPT-6 Astra; Pro 5x ($100) multiplies those by about five and Pro 20x ($200) by about twenty. A weekly limit may also apply on top, cloud chats (which run Sol) can use more than local messages, and a large multi-file task consumes far more of the allowance than a quick question.
Is there a Codex weekly limit?
Yes, potentially. OpenAI's pricing page says local messages and cloud chats share the plan's allowance and that weekly limits may also apply, without publishing the weekly number. In practice the five-hour window is the one you feel first on an ordinary day; the weekly cap is the one you feel after several heavy days in a row. The usage dashboard in ChatGPT and the /status command in the Codex CLI show your current limits and reset times.
What happens when Codex hits the usage limit mid-task?
The agent is allowed to finish the turn it is on, subject to fair use, so a long refactor does not die halfway. After that you wait for the window to reset, switch to a smaller model to stretch what is left, buy additional credits (available on Plus and Pro without changing plan), or run extra local sessions against an API key billed at usage rates. Upgrading the plan is the fourth option, not the first.
How do Codex Pro 5x and Pro 20x limits compare to Plus?
Pro is sold in two tiers as of 2026: 5x ($100/month) and 20x ($200/month), each a multiple of the Plus allowance. On OpenAI's own estimates that means roughly 50–500 GPT-5.6 Sol messages per five-hour window on Pro 5x and 200–2,000 on Pro 20x, against 10–100 on Plus. On the newer GPT-6 Sol the same ladder reads 15–150, 70–700 and 300–3,000 (OpenAI's page, read September 28, 2026). A standard Business seat ($20 per user per month billed annually, $25 monthly) gets the Plus estimates; the $100 Business seat gets the Pro 5x ones. New sign-ups to Pro 20x have been paused since September 10, 2026.
Why did I hit the Codex limit so fast?
Because the allowance is spent by tokens, not messages. Model choice, context size, reasoning depth, tool use, retrieval, and caching all change how much one message costs; cloud chats run the heaviest model by default; fast mode consumes credits at a higher rate; image generation uses the allowance roughly three to five times faster; and a session that holds a large codebase in context burns more per turn than a fresh one. Two tasks that look identical can cost very different amounts of the window.
Are Codex usage limits the same as Codex rate limits?
People use the phrases interchangeably, and OpenAI's own page mixes them. The rolling five-hour allowance and the weekly cap are what most people mean by either phrase. A separate meaning of rate limit applies to API keys, where requests and tokens per minute are governed by your API tier rather than a ChatGPT plan; if you run Codex against an API key, those limits and usage-based billing apply instead of the plan windows.
What is the Codex rate limit?
On a ChatGPT plan, the Codex rate limit is the plan allowance: a per-model estimate of local messages per five-hour window (OpenAI's page, read September 15, 2026, gives roughly 10 to 100 GPT-5.6 Sol messages on Plus, 50 to 500 on Pro 5x, and 200 to 2,000 on Pro 20x) plus a weekly limit that OpenAI says may apply without publishing the number. Both reset on schedules shown in the ChatGPT usage dashboard and via /status and /usage in the Codex CLI. The other meaning, API rate limits in requests and tokens per minute, applies only when you run Codex with an API key, and that route is billed at usage rates instead of the plan windows.

More on Vibe Code Rescue

Cursor Usage Limits (2026): Why Every Tier Is a Multiple of a Number Cursor Never Publishes

Cursor's pricing page sells Pro at $20, Pro+ at $60 and Ultra at $200, and describes what you get on each as "extended", "3x Pro" and "20x Pro" agent limits. It never says what Pro's limit is. Here is what is actually published as of September 2026, the one piece of arithmetic those multipliers do allow, and why Pro+ is the tier that buys you nothing per dollar.

6 min

Base44 Tutorial (2026): Build and Ship a Working Internal Tool in an Afternoon, Step by Step

A Base44 tutorial from someone who maintains 8+ Base44 apps: build a job tracker with a form, a table, and a status filter in ten prompts, add login and permissions, import your real data, and publish it, on the free plan's 25 credits. The exact prompt at each step, the four places builds go wrong, and the point where you stop and rebuild properly.

9 min

Is Emergent Free? (2026) What the $0 Plan's 10 Credits Actually Buy, Where the $20 Tier Starts, and the Cost That Is Not on the Pricing Page

Emergent has a free plan, and it is exactly as free as ten credits a month. Read from emergent.sh/pricing on September 15, 2026: Free is $0 with 10 monthly credits, Standard is $20 a month ($17 annual) for 100 credits and private hosting, Pro is $200 ($167 annual) for 750 credits, Enterprise is custom. Here is what ten credits buys on an agent that plans, codes, and deploys a whole app, the moment you have to pay, how it lines up against Lovable's and Base44's free tiers, and the cost every free plan leaves off the page.

5 min

Antigravity vs Claude Code (2026): Google's $0 Agent IDE That Runs Claude vs the Terminal Agent Metered on Your Claude Plan

Antigravity vs Claude Code, as of September 2026: Google's Antigravity is generally available at $0 for individuals with unlimited Tab and Command and weekly rate limits on its agents, and its model menu includes Claude Sonnet and Opus 4.6. Claude Code is Anthropic's terminal agent, included on Pro from $17 a month, Max from $100, and Team seats, running the current Sonnet 5 and Opus 5 against a five-hour usage pool. Free last-generation Claude in an IDE vs current Claude in a terminal: the decision, the limits, and the math.

6 min
Justin McKelvey, Fractional CTO and AI consultant in Austin, TX

Written by

Justin McKelvey

Fractional CTO & AI consultant in Austin, TX. 15 years building software, 50+ products shipped, $53M+ in client revenue generated. I help $1M–$50M founders ship production software and automate operations with AI — without hiring a full-time executive team.

Work with me