Limits explained· Facts verified 2026-09-03

Codex Usage Limits: Plus vs Pro, Explained

Quick answer

Codex usage limits work on rolling 5-hour windows, shared across every Codex surface (CLI, IDE, web, iOS) on one plan allowance. On ChatGPT Plus ($20/mo) OpenAI's own estimates run from 3–30 messages per 5 hours on the top model (GPT-6 Astra) up to 250–2,000 on the lightest (GPT-5.6 Luna); Pro ($100–200/mo) is 5x or 20x that. When you see 'You've hit your usage limit. Try again later.', your options are: wait for the window, switch to a lighter model, buy credits, upgrade, or move routine work to another provider.

How the limits actually work

When you hit the limit, in order of cheapness

  1. Switch models down for routine work — the biggest free lever.
  2. Wait: check the usage dashboard; the rolling window may free capacity within the hour.
  3. Buy credits for a genuine deadline crunch.
  4. Upgrade Plus → Pro ($100 for 5x) only if you cap out most days — occasional capping is cheaper solved with credits.
  5. Offload volume work: point routine agent tasks at a different provider and save Codex quota for what needs it.

The 2026 plan ladder

PlanPriceCodex allowance
Free$0Limited Codex access
Go$8/moEntry allowance
Plus$20/moStandard allowance — the community default
Pro$100–200/mo5x or 20x Plus
Business$20/user/mo (annual)Plus-level per seat

From OpenAI's official pricing docs, 2026-09-03. API pay-as-you-go has no windows — the meter is the limit.

The honest catch

The limits are estimates, not contracts — OpenAI publishes ranges, tightened them in 2026, restored them after backlash, and can change them again. If your workflow depends on a guaranteed volume, a windowed subscription cannot promise it; only a metered API or a budgeted flat plan can.

Where Standard Compute fits

Codex CLI supports custom OpenAI-compatible providers, and about a quarter of Standard Compute's active customers run it against our API: base URL api.stdcmpt.com/v1, one flat plan from $19/mo, smart routing, no 5-hour windows. The pattern that works: keep your ChatGPT plan for what needs OpenAI's top models, run the volume work through a flat budget.

Get your API key →Compare our plans honestly →

FAQ

How long until Codex limits reset?

There is no fixed reset time — windows roll continuously over 5 hours, so capacity frees up as older usage ages out. The usage dashboard shows where you stand. Pro plans also reference weekly caps for sustained heavy use.

Why did I hit the limit after a few messages?

Almost always the model: top-tier models like GPT-6 Astra allow as few as 3–30 messages per 5-hour window on Plus. Agent runs also consume multiple requests per visible message — one 'message' of agentic work can be dozens of calls.

Is Codex Pro worth $200?

The 20x tier only pays off for parallel, sustained agentic work — the same logic as Claude Max 20x. Most heavy users land on the $100 5x tier, and community experience says even that mostly matters on the top models.

Related money questions

Is Claude Max Worth It? ($100 vs $200)Plan reviewHow Much Does OpenCode Actually Cost?Cost guideCopilot AI Credits Ran Out — What Changed and What to DoLimits explained