Justin McKelvey

Justin McKelvey

Fractional CTO · 15 years, 50+ products shipped

• AI for Business • 7 min read •

GPT-6 Astra (2026): What It Costs, Who Actually Gets It, and GPT-6 vs Claude Opus 5 for a Business

GPT-6 Astra: the short answer

As of September 2026, GPT-6 Astra is OpenAI's top model, and it costs $10 per million input tokens and $50 per million output on the API: exactly twice Claude Opus 5, and the same as Claude Fable 5.1. It has a 1,050,000-token window, a price line at 272K input tokens that reprices the whole request, and on ChatGPT Plus it runs 5 to 45 Codex messages per five hours. Every new flagship launch produces the same two weeks of benchmark screenshots and "this changes everything" threads. It changes one thing for most businesses: the price of the hardest ten percent of your AI work. This page is the price, the catch, who actually gets it, and what an owner-led company should do, with every number read from OpenAI's and Anthropic's own pages on September 11, 2026.

What GPT-6 Astra costs on the API

From OpenAI's GPT-6 Astra model page: $10 per million input tokens, $1 per million cached input tokens, $12.50 per million cache writes, and $50 per million output tokens. Batch and Flex processing are priced at 50% of those rates ($5 in, $25 out), and Fast mode is priced at 2x. The model takes text and images, returns text, supports reasoning effort from low to max, and lists web search, file search, code interpreter, computer use, and MCP among its tools. Knowledge cutoff: April 30, 2026.

For context, the rest of OpenAI's current lineup on the same pricing page, as of September 27, 2026: the two newer GPT-6 models, GPT-6 Sol at $2 / $10 and GPT-6 Luna at $0.10 / $0.50 (both released September 22), and the GPT-5.6 family, still listed at Sol $4 / $20, Terra $2 / $12 and Luna $0.20 / $1.20. Astra is 5x GPT-6 Sol and 100x GPT-6 Luna on input. The full Claude-versus-OpenAI API comparison, beyond the flagships, is on Claude API vs OpenAI API.

The 272K line: the math OpenAI's pricing page does not do for you

The window is 1,050,000 tokens (922,000 input, 128,000 output). The line that matters sits much lower. OpenAI's words: "Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request." Not the tokens over the line. The full request.

Run one document review both ways. A 270,000-token contract bundle with a 10,000-token answer: 270K at $10 is $2.70, plus 10K at $50 is $0.50, so $3.20. Add 30,000 tokens of exhibits and the same review is 300K at $20 ($6.00) plus 10K at $75 ($0.75), so $6.75. Eleven percent more input, a bill more than twice as large. The same 270K job on Claude Opus 5 at Anthropic's listed $5 / $25 is $1.60.

If you build anything that stuffs whole folders into a prompt, the first optimization is not the model. It is keeping each request under 272K, which usually means sending the three relevant documents instead of the thirty.

GPT-6 in ChatGPT and Codex: who actually gets it

OpenAI announced Astra on September 3, 2026, as reported by launch coverage, and rolled it out in stages: a limited set of organizations first, then ChatGPT Plus, Pro, Business, and Enterprise, plus the API. OpenAI's Codex pricing page is the clearest first-party map of who gets how much. Its estimates for GPT-6 Astra, in local messages per five-hour window: Plus ($20) 5 to 45, Pro 5x ($100) 25 to 225, Pro 20x ($200) 100 to 900, and Business the same as Plus. The Free and Go plans have no Astra row. Those are estimates, not caps; a big multi-file task uses far more of the window than a quick question, and weekly limits may apply on top. The mechanics of that window are on Codex usage limits, and the plan ladder itself is on Codex pricing.

Past the included limits, Codex bills in credits, and Astra is the expensive one: 250 credits per million input tokens and 1,250 per million output, with Fast mode in Codex at 2.5x Astra's standard rate. Five to 45 messages is a taste, not a workday. On Plus, Astra is the model you save for the hard problem and not the one you leave selected.

GPT-6 Astra vs GPT-5.6 Sol, Terra, and Luna

Correction (September 27, 2026): when this page went up, Astra was the only GPT-6 model. On September 22 OpenAI added GPT-6 Sol ($2 / $10 per million tokens), which its model page says is built "to power complex coding and agentic workflows," and GPT-6 Luna ($0.10 / $0.50), its "most efficient model for focused, high-volume tasks." Both carry the same 272K pricing line as Astra. GPT-6 Sol is half the price of GPT-5.6 Sol and undercuts Terra on output; GPT-6 Luna is half of GPT-5.6 Luna. The GPT-5.6 list below still describes models OpenAI sells at those prices, but for new API work, the GPT-6 trio (Astra, Sol, Luna) is the ladder to start from. Every rate is on OpenAI API pricing.

  • Astra ($10 / $50): the hardest end-to-end work, in OpenAI's own description. Long research, complex code changes, computer use.
  • Sol ($4 / $20): OpenAI's pick for complex reasoning and high-stakes decisions, and the model Codex cloud chats run on.
  • Terra ($2 / $12): the everyday workhorse: production tasks, reporting, document analysis.
  • Luna ($0.20 / $1.20): fast, high-volume work such as routing, classification, extraction, and support.

The practical split I set up: Luna or Terra for anything that runs a thousand times a day, Sol for the judgment calls, Astra only when a wrong answer costs more than the tokens.

GPT-6 vs Claude Opus 5 (and Fable 5.1) on price

One change since the original read: on September 22, 2026 Anthropic launched Claude Opus 5.5 at $4 / $20, with Opus 5 still listed at $5 / $25, so against the newest Opus, Astra is 2.5x rather than 2x. Anthropic's pricing page on September 11, 2026: Claude Opus 5 at $5 / $25, Claude Sonnet 5 at $2 / $10, Claude Haiku 4.5 at $1 / $5, and Claude Fable 5.1 at $10 / $50, with cache hits at a tenth of the input rate. So at the top, Astra matches Fable token for token and costs twice Opus 5. One level down, Sonnet 5 and Terra are both $2 in, with Terra's output $2 higher. At the bottom, OpenAI wins outright: Luna is a fifth of Haiku 4.5's price. The per-model Anthropic rates, caching, and batch discounts are broken down on Anthropic API pricing.

Which one is better is a question for your own work, not a launch-week benchmark. Take five real tasks, run them on Astra and Opus 5, and compare the answers and the bills. If the flagship difference does not show up in your five tasks, the only place it will show up is the invoice.

Is GPT-6 free?

No. On the API every token is billed, and on OpenAI's own Codex table Astra starts at Plus. The free and $8 Go plans run the GPT-5.6 family. That is fine: the free tier's job is to get you to Plus, and Astra's job is to get Plus users to Pro.

What a business should actually do

Do not switch assistants because a new flagship launched. If your team is on ChatGPT Plus or Business, Astra shows up inside the plan you already pay for; tell people it exists and that it is for the hard problems. If your team is on Claude, Opus 5 does the same job at half the API price, and nothing about this launch is a reason to migrate a working setup. If you build on the API, keep requests under 272K, route volume to the cheap models, and give the flagship the few jobs that earn it. And if what you want is the assistant, the connectors, and the first three workflows installed rather than benchmarked, the Claude for Small Business install does that in about two weeks, from $4,500; the $397 Playbook is the do-it-yourself version. The model is the part that changes every quarter. The workflow is the part that pays.

Update: the $200 Pro pause, and the $500 tier after it

Update (September 20, 2026): OpenAI paused new sign-ups and upgrades to Pro 20x ($200/month) on September 10, 2026, after demand for GPT-6 Astra (launched September 3) outran capacity. As of September 29 the pause is lifted: OpenAI's help center says Pro 200 is open to new subscriptions again, though anyone not grandfathered gets a lower usage allowance, and the new Pro 500 at $500/month is the only Pro tier with Astra Ultrafast. For anyone moving up from Plus to get more Astra, the first step is still Pro at $100.

Free Resource Justin McKelvey

Get the Free AI Content Toolkit

The exact system I use to turn one idea into a month of content — atomization framework, voice template, prompt library, weekly system.

Frequently Asked Questions

How much does GPT-6 cost?
As of September 2026, GPT-6 Astra costs $10 per million input tokens, $1 per million cached input tokens, and $50 per million output tokens on OpenAI's API, per OpenAI's model page. Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the whole request ($20 and $75). Batch and Flex are half price ($5 and $25) and Fast mode is double. In ChatGPT it comes with the paid plans rather than a separate price: Plus is $20 a month, Pro starts at $100.
Is GPT-6 free?
Not in any way that matters for work. OpenAI's Codex pricing page lists GPT-6 Astra usage for Plus ($20), Pro 5x ($100), Pro 20x ($200), and Business, and has no Astra row for the Free or Go plans. On the API every token is billed. Astra is the model you pay for.
When was GPT-6 released?
OpenAI announced GPT-6 Astra on September 3, 2026, as reported by launch coverage, with a staged rollout: a limited set of organizations first, then ChatGPT Plus, Pro, Business, and Enterprise over the following days, plus the API. OpenAI's API documentation lists the model as gpt-6-astra with an April 30, 2026 knowledge cutoff.
GPT-6 vs Claude Opus 5: which is better for a business?
On price, Claude Opus 5 is half the cost: $5 per million input tokens and $25 output against GPT-6 Astra's $10 and $50, per both vendors' pricing pages in September 2026. On capability they are both flagship models aimed at the hardest work, and the honest answer is that most business tasks do not need either one; Claude Sonnet 5 ($2 / $10) or GPT-5.6 Terra ($2 / $12) handle the daily volume. Use the flagship for the few jobs where a better answer is worth paying double, and test both on your own work before you commit. Since September 22, 2026, Anthropic's newer Claude Opus 5.5 is $4 / $20, so Astra is now 2.5x the current Opus on the API.
What is GPT-6's context window?
GPT-6 Astra has a 1,050,000-token context window, with up to 922,000 input tokens and 128,000 output tokens per request, per OpenAI's model page. The catch is the price line at 272K: any prompt with more than 272K input tokens is billed at 2x the input and cache rates and 1.5x the output rate for the entire request, not just the tokens over the line.
Is GPT-6 Astra worth it for a small business?
For the company-wide assistant, it is not a reason to change anything: if your team is on ChatGPT Plus or Business, Astra arrives inside the plan you already pay for, with 5 to 45 Codex messages per five hours on the $20 tier. For anything you build on the API, it is worth it only for the small set of tasks where the best available answer pays for a 2x token bill. Route the volume work to a cheaper model (GPT-6 Luna at $0.10 / $0.50 since September 22, or Claude Haiku 4.5 at $1 / $5), put the everyday work on GPT-6 Sol at $2 / $10, and keep the flagship for the hard cases.

More on AI for Business

ChatGPT Business vs Plus (2026): Which One a Small Business Should Actually Pay For

ChatGPT Plus vs Business, as of September 2026: Plus is $20 a month for one person and OpenAI may train on your chats unless you opt out; Business is $20 a seat billed annually or $25 monthly, two seats minimum, with no training by default, admin controls, shared projects for groups, and the Company Knowledge plugin. The side-by-side from OpenAI's pages, and three owner scenarios with a verdict each.

6 min

ChatGPT for Customer Service (2026): What Works, What Breaks, and What It Costs a Small Business

ChatGPT for customer service, as of September 2026: it drafts replies, answers policy questions from your own documents, and triages an inbox for $20 to $25 a seat a month on ChatGPT Business. It can't look up an order, issue a refund, or remember a customer unless you connect the system that holds them. The honest capability line, the plan you need, a setup in six steps, the failure modes, and the point where a real agent takes over.

7 min

ChatGPT Pulse (2026): What It Was, Why OpenAI Retired It, and How to Get the Morning Brief for Your Business

ChatGPT Pulse, the daily research cards OpenAI previewed for Pro users in September 2025, was sunset on June 17, 2026 with a 14-day wind-down. What it did, which plans ever had it, why scheduled tasks replaced it, and how an owner rebuilds the morning brief on a Plus or Business seat without handing ChatGPT the wrong things.

7 min

TRAIGA: What the Texas Responsible AI Governance Act Actually Requires of a Business (2026)

TRAIGA, the Texas Responsible AI Governance Act, took effect January 1, 2026. It applies to any business that operates in Texas, with no size exemption, but most of its duties bind state agencies. For a private company it is a short list of things you may not build AI to do on purpose, one disclosure duty for health care, and fines that start at $10,000. Here is the bill text in plain English.

9 min
Justin McKelvey, Fractional CTO and AI consultant in Austin, TX

Written by

Justin McKelvey

Fractional CTO & AI consultant in Austin, TX. 15 years building software, 50+ products shipped, $53M+ in client revenue generated. I help $1M–$50M founders ship production software and automate operations with AI — without hiring a full-time executive team.

Work with me

If this was useful, here are two ways I can help: