Justin McKelvey
Fractional CTO · 15 years, 50+ products shipped
Gemini Usage Limits (2026): What Free, AI Pro, and AI Ultra Actually Cap, and When They Reset
Gemini usage limits: the short answer
As of September 2026, Google publishes no prompt count for any Gemini plan. What it publishes is a compute-based limit that refreshes every 5 hours until you hit a weekly cap, and a set of multipliers: the free tier gets "standard limits," Google AI Plus ($4.99) gets 2x, Google AI Pro ($19.99) gets 4x, and Google AI Ultra gets 5x AI Pro at $99.99 or 20x at $199.99. Media generation, Deep Research, the Pro model, and Deep Think burn the allowance fastest, and paid users who hit the wall can keep chatting on Flash-Lite. Everything below comes from Google's Gemini subscriptions page and the Gemini Apps help center, read on September 15, 2026.
If you came here looking for "50 prompts a day," that number stopped existing on May 17, 2026, when Google moved Gemini to compute-based metering. Here's what replaced it. The same question for every other vendor, side by side, is on AI usage limits by plan.
Gemini limits on Free, AI Plus, AI Pro and AI Ultra
| Plan | Price | Usage limit | Context window | What's included | Window and reset |
|---|---|---|---|---|---|
| Free | $0 | Standard limits | 32K tokens | 3.6 Flash, "varying access" to 3.1 Pro, image generation, Deep Research, Gemini Live, Canvas, Gems | Refreshes every 5 hours until the weekly limit |
| Google AI Plus | $4.99/month | 2x Free | 128K tokens | Everything in Free plus video generation, Daily Brief, 200 Flow credits, 400 GB storage | Refreshes every 5 hours until the weekly limit |
| Google AI Pro | $19.99/month | 4x Free | 1M tokens | Everything in Plus, 1,000 Flow credits, Gemini Spark, higher Jules limits, 5 TB storage | Refreshes every 5 hours until the weekly limit |
| Google AI Ultra 5x | $99.99/month | 5x AI Pro | 1M tokens | Everything in Pro, Deep Think, 10,000 Flow credits, 20 TB storage | Refreshes every 5 hours until the weekly limit |
| Google AI Ultra 20x | $199.99/month | 20x AI Pro | 1M tokens | Everything in Ultra 5x, 25,000 Flow credits, 30 TB storage | Refreshes every 5 hours until the weekly limit |
| Workspace Business Starter / Standard / Plus | $7 / $14 / $22 per user/month, annual | Not published | Not published | Gemini in Gmail (Starter); Gmail, Docs, Meet and expanded Gemini app access (Standard and up) | Not published |
| Gemini Enterprise, Business edition | From $21/seat/month | Not published | Not published | Gemini Enterprise app, connectors, no-code agents, up to 300 seats | Not published |
Two things the table can't show. First, the multipliers are relative to a "standard limit" Google never states in absolute terms, so 4x is a real ratio applied to an unknown base. Second, all three models (Flash-Lite, Flash, and Pro) are available on every tier including Free; what the paid tiers buy is more compute against them, not access to them. My verdict on whether the $19.99 tier earns its keep is on is Gemini worth it.
The per-unit math Google's ladder hides
The unknown base cancels out if you price everything in it. Call the free tier's allowance one unit. AI Ultra is sold as a multiple of AI Pro, and AI Pro is 4x Free, so Ultra 5x is 20 units and Ultra 20x is 80.
| Plan | Price a month | Free-tier units | Price per unit |
|---|---|---|---|
| Google AI Plus | $4.99 | 2 | $2.50 |
| Google AI Pro | $19.99 | 4 | $5.00 |
| Google AI Ultra 5x | $99.99 | 20 | $5.00 |
| Google AI Ultra 20x | $199.99 | 80 | $2.50 |
So the cheapest Gemini, by allowance, is the $4.99 tier, tied with the $199.99 one. AI Pro and Ultra 5x charge twice as much per unit. That doesn't make AI Pro a bad buy, because the $15 gap mostly pays for other things: the 1M context window instead of 128K, 1,000 Flow credits instead of 200, and 5 TB of storage instead of 400 GB. It does mean that if your only complaint on Free is running out, AI Plus is the upgrade that fixes exactly that at the lowest price per unit. The assumption: Google's multipliers apply to the whole allowance, the 5-hour refresh and the weekly cap alike, which its table implies but doesn't spell out.
How the Gemini meter works
Google's help center describes the limit in one sentence: "compute-based usage limits that determine how much you can interact with Gemini tools and features. These limits factor in the complexity of your prompt, the models and features you use, and the length of your chat." Three consequences follow.
- The model you pick sets the burn rate. Flash-Lite is the "efficient workhorse," Flash balances speed and reasoning, and Pro is "our most advanced model" with longer response times. Google says outright that "more advanced models and higher thinking levels consume more of your usage."
- Chat length compounds. A long thread costs more per turn than a short one, because the whole conversation is context. Starting a new chat for a new topic is the cheapest habit on this page.
- Features are priced in compute too. Google lists the expensive ones: image, video, and music generation, Deep Research, the Pro model, and extended thinking or Deep Think. Deep Think is Ultra-only and requires the Pro model, so it's the most expensive thing you can ask Gemini to do.
You can see where you stand at gemini.google.com under Settings, then Usage Limits. Gemini also notifies you when you're close, and again when you hit it, with the refresh time.
The 5-hour refresh and the weekly cap
Gemini uses the same two-layer structure OpenAI and Anthropic use, and Google's wording is the clearest of the three: "Your limit refreshes every 5 hours until you reach your weekly limit." So there are two ceilings. The 5-hour one refills on its own during the day. The weekly one sits above it, and once you've spent that, the 5-hour refresh stops helping until the week rolls over. Google doesn't publish how large the weekly limit is on any tier, or exactly when it resets, so I'm not going to invent either.
The practical read: a Pro subscriber who runs three Deep Research reports and a video generation before lunch can hit the 5-hour ceiling, wait, and be fine by mid-afternoon. A subscriber who does that every day will meet the weekly ceiling instead, and that one doesn't refresh at 3pm.
Deep Research, video, and images
None of these carries a published count. Google's feature table shows Deep Research available on every tier including Free, video generation on the paid tiers, and image generation with Nano Banana 2 everywhere, but the only quantitative statement is that they "consume more of your usage." Two tier-specific notes do appear. For users without a Google AI plan, "some features that require more compute resources, like Deep Research, may be unavailable during periods of high demand." And when capacity changes, "limits for users without a Google AI plan may be limited before users with a plan." Free is first in line to lose Deep Research on a busy day.
Video and image creation outside the Gemini app, in Google Flow, run on a separate credit system that Google does quantify: 200 Flow credits on Plus, 1,000 on Pro, 10,000 or 25,000 on Ultra. That's the one place in Google's plan lineup where a hard number exists.
The context window is the published number
Google is precise about the thing it can measure. The free tier's context window is 32K tokens, AI Plus is 128K, and AI Pro and Ultra are 1M, which Google translates as "up to 1,500 pages of text or 30,000 lines of code." Exceeding the window doesn't stop the chat; it means Gemini stops seeing all of what you gave it, and Google warns responses may "miss connections or details throughout the content." For anyone uploading a contract or a codebase, the jump from 32K to 1M is the real reason to pay $19.99, more than the 4x multiplier.
Google only gives the pages figure for 1M. Scaling it down in proportion, which is my arithmetic and not Google's, gives a rough guide to which plan your documents need:
| Plan | Context window | Roughly this much text |
|---|---|---|
| Free | 32K tokens | about 48 pages |
| Google AI Plus | 128K tokens | about 190 pages |
| Google AI Pro and Ultra | 1M tokens | up to 1,500 pages (Google's figure) |
On that scale a 60-page contract is already more than Free can hold at once, and a 300-page document set needs Pro.
Hit the Gemini limit? Flash-Lite, AI credits, or wait
- Subscribers: "you can continue your conversation with Flash-Lite." The chat keeps going on the cheapest model until your limit refreshes.
- Anyone: upgrade to a higher plan, or wait for the refresh time Gemini shows you.
- AI credits: the subscriptions page says you "can extend your limits by purchasing AI credits." Google doesn't publish the price per credit on the pages I read.
- Limits can move. Google says "limits may change without notice, including due to capacity constraints." Plan for the multiplier, not for a fixed quota.
If you keep hitting the app's caps for work that runs the same way every time, the API bills per token instead of resetting on the app's five-hour window (it has its own rate limits). Gemini 3.8 Flash is $0.75 in and $3.75 out per million tokens until December 31, and the free tier's catch is that Google may use what you send to improve its products. The numbers are on the Gemini API pricing page.
Gemini vs ChatGPT on limits
The structures are close and the disclosure isn't. ChatGPT treats plain chat as unlimited on every plan and meters only reasoning, GPT-6 Pro and its agent work, and for those OpenAI prints actual per-window ranges. Gemini meters everything, chat included, and Google prints only ratios. If you value knowing your number, ChatGPT tells you more; if you value a $4.99 tier and a 1M context window at $19.99, Gemini is the cheaper ladder. The full ChatGPT side is on ChatGPT usage limits, and the head-to-head on everything else is on Gemini vs ChatGPT.
Five habits that stretch a Gemini allowance
- Default to Flash. Reserve Pro for the tasks that need reasoning across a long document, and Deep Think for almost nothing.
- New chat, new topic. The single biggest lever, since chat length is one of the three inputs Google names.
- Batch Deep Research. Run reports in one block early in a 5-hour window, so the refresh lands before you need the next set.
- Watch the weekly, not the 5-hour. Check Settings, then Usage Limits mid-week. The 5-hour ceiling is a pause; the weekly one is a stop.
- Business accounts: don't assume the multipliers. Workspace and Gemini Enterprise publish no limit. Test with one seat for a month before rolling it out.
Sources
- Google, Gemini subscriptions (plan prices, 2x / 4x / 5x / 20x, the compute-based footnote, Flow credits, storage, 1,500 pages): read September 15, 2026
- Gemini Apps Help, limits and upgrades for Google AI subscribers (usage table, context windows, Flash-Lite fallback, Settings > Usage Limits, capacity wording): read September 15, 2026
- Gemini Apps Help, compute-based usage limits (the May 17, 2026 change, which features burn faster): read September 15, 2026
- Gemini Apps Help, continuing with Flash-Lite: read September 15, 2026
- Google Workspace pricing (Business Starter, Standard, Plus; Gemini Enterprise Business edition): read September 15, 2026
Hit the cap? Every plan's limit, one page.
ChatGPT, Codex, Claude, Gemini, and Copilot: what's metered, what resets when, and the cheapest way around each. Re-sent when a limit moves.
Frequently Asked Questions
- What are Gemini's usage limits?
- As of September 2026, Gemini uses compute-based limits that refresh every 5 hours until you reach a weekly limit, and Google publishes the tiers as multipliers rather than counts: standard limits on the free tier, 2x on Google AI Plus ($4.99), 4x on Google AI Pro ($19.99), and 5x or 20x AI Pro on Google AI Ultra ($99.99 or $199.99). How fast you use the allowance depends on prompt complexity, the model, the features, and how long the chat is.
- How many prompts a day do you get with Gemini?
- Google doesn't publish a prompts-per-day number for any plan, and since May 17, 2026 the meter isn't a prompt count at all. It's compute: a short question on Flash costs a fraction of what a Deep Research run or a video generation costs. The only quantities Google prints are the multipliers (2x, 4x, 5x, 20x) and the context windows (32K tokens free, 128K on Plus, 1M on Pro and Ultra).
- What happens when you hit the Gemini usage limit?
- Gemini warns you as you get close, then tells you when the limit refreshes. If you have a Google AI subscription you can keep the conversation going on Flash-Lite, Google's fastest and cheapest model, or buy AI credits to extend the limit. Free users wait for the 5-hour refresh, and during high demand Google says compute-heavy features like Deep Research may be unavailable to them first.
- Does Google AI Pro have unlimited Gemini usage?
- No. Google AI Pro at $19.99 a month gives you 4x the free tier's limit, a 1M token context window, and access to features like Deep Research and video generation, but the same 5-hour refresh and weekly cap apply. Google says limits may change without notice, including for capacity reasons, and that paid users are limited after free users when that happens.
- What uses up the Gemini allowance fastest?
- Google lists the expensive features by name: image, video, and music generation, Deep Research, the Pro model, and extended thinking or Deep Think (Ultra only). Its help center also says more advanced models and higher thinking levels consume more of your usage, and that longer chats cost more per turn. Flash-Lite for routine questions and a fresh chat for each new topic are the two cheapest habits.
- Is the Gemini limit different on a Workspace business account?
- Google's Workspace pricing page lists Gemini in the Business Starter ($7), Standard ($14), and Plus ($22 per user a month, billed annually) tiers, with Standard and above getting expanded access to models and features, but it publishes no usage numbers for any of them. Gemini Enterprise's Business edition starts at $21 a seat a month and also publishes no limit. The multiplier table above applies to personal Google Accounts only.
More on AI for Business
ChatGPT Business vs Plus (2026): Which One a Small Business Should Actually Pay For
ChatGPT Plus vs Business, as of September 2026: Plus is $20 a month for one person and OpenAI may train on your chats unless you opt out; Business is $20 a seat billed annually or $25 monthly, two seats minimum, with no training by default, admin controls, shared projects for groups, and the Company Knowledge plugin. The side-by-side from OpenAI's pages, and three owner scenarios with a verdict each.
ChatGPT for Customer Service (2026): What Works, What Breaks, and What It Costs a Small Business
ChatGPT for customer service, as of September 2026: it drafts replies, answers policy questions from your own documents, and triages an inbox for $20 to $25 a seat a month on ChatGPT Business. It can't look up an order, issue a refund, or remember a customer unless you connect the system that holds them. The honest capability line, the plan you need, a setup in six steps, the failure modes, and the point where a real agent takes over.
ChatGPT Pulse (2026): What It Was, Why OpenAI Retired It, and How to Get the Morning Brief for Your Business
ChatGPT Pulse, the daily research cards OpenAI previewed for Pro users in September 2025, was sunset on June 17, 2026 with a 14-day wind-down. What it did, which plans ever had it, why scheduled tasks replaced it, and how an owner rebuilds the morning brief on a Plus or Business seat without handing ChatGPT the wrong things.
AI for MSPs: The Permissions-First AI Work Small-Business Clients Will Actually Pay For (2026)
Your clients don't need a chatbot from their MSP. They need the AI they already bought to stop showing people files they shouldn't see, and to get used. Here's the 30-day job, the Microsoft licensing detail that makes it cheap to start, and what not to sell.
Written by
Justin McKelvey
Fractional CTO & AI consultant in Austin, TX. 15 years building software, 50+ products shipped, $53M+ in client revenue generated. I help $1M–$50M founders ship production software and automate operations with AI, without hiring a full-time executive team.
Work with meIf this was useful, here are two ways I can help:
Before you go
Every plan's limits on one page.
Free: what each plan actually caps, what resets when, and the cheapest way around each one. 25 plans, re-sent when a limit moves.
Send me the cheat sheet →