Justin McKelvey
Fractional CTO · 15 years, 50+ products shipped
GPT-6 Astra (2026): What It Costs, Who Actually Gets It, and GPT-6 vs Claude Opus 5 for a Business
GPT-6 Astra: the short answer
As of September 2026, GPT-6 Astra is OpenAI's top model, and it costs $10 per million input tokens and $50 per million output on the API: exactly twice Claude Opus 5, and the same as Claude Fable 5.1. It has a 1,050,000-token window, a price line at 272K input tokens that reprices the whole request, and on ChatGPT Plus it runs 5 to 45 Codex messages per five hours. Every new flagship launch produces the same two weeks of benchmark screenshots and "this changes everything" threads. It changes one thing for most businesses: the price of the hardest ten percent of your AI work. This page is the price, the catch, who actually gets it, and what an owner-led company should do, with every number read from OpenAI's and Anthropic's own pages on September 11, 2026.
What GPT-6 Astra costs on the API
From OpenAI's GPT-6 Astra model page: $10 per million input tokens, $1 per million cached input tokens, $12.50 per million cache writes, and $50 per million output tokens. Batch and Flex processing are priced at 50% of those rates ($5 in, $25 out), and Fast mode is priced at 2x. The model takes text and images, returns text, supports reasoning effort from low to max, and lists web search, file search, code interpreter, computer use, and MCP among its tools. Knowledge cutoff: April 30, 2026.
For context, the rest of OpenAI's current lineup on the same pricing page: GPT-5.6 Sol at $4 / $20, GPT-5.6 Terra at $2 / $12, and GPT-5.6 Luna at $0.20 / $1.20. Astra is 2.5x Sol and 50x Luna on input. The full Claude-versus-OpenAI API comparison, beyond the flagships, is on Claude API vs OpenAI API.
The 272K line: the math OpenAI's pricing page does not do for you
The window is 1,050,000 tokens (922,000 input, 128,000 output). The line that matters sits much lower. OpenAI's words: "Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request." Not the tokens over the line. The full request.
Run one document review both ways. A 270,000-token contract bundle with a 10,000-token answer: 270K at $10 is $2.70, plus 10K at $50 is $0.50, so $3.20. Add 30,000 tokens of exhibits and the same review is 300K at $20 ($6.00) plus 10K at $75 ($0.75), so $6.75. Eleven percent more input, a bill more than twice as large. The same 270K job on Claude Opus 5 at Anthropic's listed $5 / $25 is $1.60.
If you build anything that stuffs whole folders into a prompt, the first optimization is not the model. It is keeping each request under 272K, which usually means sending the three relevant documents instead of the thirty.
GPT-6 in ChatGPT and Codex: who actually gets it
OpenAI announced Astra on September 3, 2026, as reported by launch coverage, and rolled it out in stages: a limited set of organizations first, then ChatGPT Plus, Pro, Business, and Enterprise, plus the API. OpenAI's Codex pricing page is the clearest first-party map of who gets how much. Its estimates for GPT-6 Astra, in local messages per five-hour window: Plus ($20) 5 to 45, Pro 5x ($100) 25 to 225, Pro 20x ($200) 100 to 900, and Business the same as Plus. The Free and Go plans have no Astra row. Those are estimates, not caps; a big multi-file task uses far more of the window than a quick question, and weekly limits may apply on top. The mechanics of that window are on Codex usage limits, and the plan ladder itself is on Codex pricing.
Past the included limits, Codex bills in credits, and Astra is the expensive one: 250 credits per million input tokens and 1,250 per million output, with Fast mode in Codex at 2.5x Astra's standard rate. Five to 45 messages is a taste, not a workday. On Plus, Astra is the model you save for the hard problem and not the one you leave selected.
GPT-6 Astra vs GPT-5.6 Sol, Terra, and Luna
- Astra ($10 / $50): the hardest end-to-end work, in OpenAI's own description. Long research, complex code changes, computer use.
- Sol ($4 / $20): OpenAI's pick for complex reasoning and high-stakes decisions, and the model Codex cloud chats run on.
- Terra ($2 / $12): the everyday workhorse: production tasks, reporting, document analysis.
- Luna ($0.20 / $1.20): fast, high-volume work such as routing, classification, extraction, and support.
The practical split I set up: Luna or Terra for anything that runs a thousand times a day, Sol for the judgment calls, Astra only when a wrong answer costs more than the tokens.
GPT-6 vs Claude Opus 5 (and Fable 5.1) on price
Anthropic's pricing page on September 11, 2026: Claude Opus 5 at $5 / $25, Claude Sonnet 5 at $2 / $10, Claude Haiku 4.5 at $1 / $5, and Claude Fable 5.1 at $10 / $50, with cache hits at a tenth of the input rate. So at the top, Astra matches Fable token for token and costs twice Opus 5. One level down, Sonnet 5 and Terra are both $2 in, with Terra's output $2 higher. At the bottom, OpenAI wins outright: Luna is a fifth of Haiku 4.5's price. The per-model Anthropic rates, caching, and batch discounts are broken down on Anthropic API pricing.
Which one is better is a question for your own work, not a launch-week benchmark. Take five real tasks, run them on Astra and Opus 5, and compare the answers and the bills. If the flagship difference does not show up in your five tasks, the only place it will show up is the invoice.
Is GPT-6 free?
No. On the API every token is billed, and on OpenAI's own Codex table Astra starts at Plus. The free and $8 Go plans run the GPT-5.6 family. That is fine: the free tier's job is to get you to Plus, and Astra's job is to get Plus users to Pro.
What a business should actually do
Do not switch assistants because a new flagship launched. If your team is on ChatGPT Plus or Business, Astra shows up inside the plan you already pay for; tell people it exists and that it is for the hard problems. If your team is on Claude, Opus 5 does the same job at half the API price, and nothing about this launch is a reason to migrate a working setup. If you build on the API, keep requests under 272K, route volume to the cheap models, and give the flagship the few jobs that earn it. And if what you want is the assistant, the connectors, and the first three workflows installed rather than benchmarked, the Claude for Small Business install does that in about two weeks, from $4,500; the $397 Playbook is the do-it-yourself version. The model is the part that changes every quarter. The workflow is the part that pays.
Get the Free AI Content Toolkit
The exact system I use to turn one idea into a month of content — atomization framework, voice template, prompt library, weekly system.
Frequently Asked Questions
- How much does GPT-6 cost?
- As of September 2026, GPT-6 Astra costs $10 per million input tokens, $1 per million cached input tokens, and $50 per million output tokens on OpenAI's API, per OpenAI's model page. Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the whole request ($20 and $75). Batch and Flex are half price ($5 and $25) and Fast mode is double. In ChatGPT it comes with the paid plans rather than a separate price: Plus is $20 a month, Pro starts at $100.
- Is GPT-6 free?
- Not in any way that matters for work. OpenAI's Codex pricing page lists GPT-6 Astra usage for Plus ($20), Pro 5x ($100), Pro 20x ($200), and Business, and has no Astra row for the Free or Go plans. On the API every token is billed. Astra is the model you pay for.
- When was GPT-6 released?
- OpenAI announced GPT-6 Astra on September 3, 2026, as reported by launch coverage, with a staged rollout: a limited set of organizations first, then ChatGPT Plus, Pro, Business, and Enterprise over the following days, plus the API. OpenAI's API documentation lists the model as gpt-6-astra with an April 30, 2026 knowledge cutoff.
- GPT-6 vs Claude Opus 5: which is better for a business?
- On price, Claude Opus 5 is half the cost: $5 per million input tokens and $25 output against GPT-6 Astra's $10 and $50, per both vendors' pricing pages in September 2026. On capability they are both flagship models aimed at the hardest work, and the honest answer is that most business tasks do not need either one; Claude Sonnet 5 ($2 / $10) or GPT-5.6 Terra ($2 / $12) handle the daily volume. Use the flagship for the few jobs where a better answer is worth paying double, and test both on your own work before you commit.
- What is GPT-6's context window?
- GPT-6 Astra has a 1,050,000-token context window, with up to 922,000 input tokens and 128,000 output tokens per request, per OpenAI's model page. The catch is the price line at 272K: any prompt with more than 272K input tokens is billed at 2x the input and cache rates and 1.5x the output rate for the entire request, not just the tokens over the line.
- Is GPT-6 Astra worth it for a small business?
- For the company-wide assistant, it is not a reason to change anything: if your team is on ChatGPT Plus or Business, Astra arrives inside the plan you already pay for, with 5 to 45 Codex messages per five hours on the $20 tier. For anything you build on the API, it is worth it only for the small set of tasks where the best available answer pays for a 2x token bill. Route the volume work to a cheaper model (GPT-5.6 Luna at $0.20 / $1.20, or Claude Haiku 4.5 at $1 / $5) and keep the flagship for the hard cases.
More on AI for Business
Gemini vs Copilot (2026): Which Is Better for a Small Business, What Each Costs, and Why the Answer Is Usually the Suite You Already Pay For
Gemini vs Copilot for a business, as of September 2026: Gemini if your company runs on Google Workspace, because it is included in the Business plans from $7 a user; Microsoft 365 Copilot if you run on Microsoft 365 and will pay $21 to $30 a seat on top of it for the version that reads your files, mail, and meetings. The consumer apps are free at both. Prices from Google's pricing page and the Microsoft canon on this site, plus the third option most owner-led companies end up choosing.
Is DeepSeek Safe? (2026): What Its Privacy Policy Actually Says, Where Your Data Goes, and the Two Ways to Use It Anyway
Is DeepSeek safe? For anything you would not post publicly, no, and the reason is in DeepSeek's own privacy policy: your prompts, device data, and IP are stored in the People's Republic of China by a Hangzhou-registered company and used to train its models unless you opt out. The model itself is a different question, because the weights are MIT-licensed and can run on a host you choose. Here is the policy, read line by line, and the rule for a business.
Kimi vs Claude (2026): The $0 Chat, the $3 Flagship API, and the Terms That Decide It for a Business
Kimi vs Claude for a business, as of September 2026: Kimi's free tier is the most generous chat in the category and its K3 flagship carries a 1M-token window at $3 in / $15 out per million tokens; Claude's Sonnet 5 is $2 / $10 with a 200k window, and its Team plan ships no-training terms and admin controls at $20 a seat. On price the two are closer than the hype says. The decision is the terms, the workspace, and what you put in the box.
Is Copilot Free? (2026): Which of the Five Copilots Costs Nothing, Which Costs $21 to $30 a Seat, and Which One Your Business Actually Wants
Is Copilot free? Two of them are. The Copilot in Windows, Edge, and the Copilot app costs nothing, and GitHub Copilot has a free tier. The one a business is usually asking about, Microsoft 365 Copilot inside Word, Excel, Outlook, and Teams, is a paid add-on: $30 a user a month on enterprise plans and, as reported, $21 on the small-business plans with an $18 promotion through September 2026. Here is the whole map, what the free one can and cannot touch, and when the $20 Claude seat is the better buy.
Written by
Justin McKelvey
Fractional CTO & AI consultant in Austin, TX. 15 years building software, 50+ products shipped, $53M+ in client revenue generated. I help $1M–$50M founders ship production software and automate operations with AI — without hiring a full-time executive team.
Work with meIf this was useful, here are two ways I can help:
Before you go
Score your AI readiness in 3 minutes.
Free, no call. 30 yes/no questions show where AI actually pays off in your business — and where it doesn't yet.
Get my readiness score →