Justin McKelvey

Justin McKelvey

Fractional CTO · 15 years, 50+ products shipped

AI for Business 6 min read

Gemini Usage Limits (2026): What Free, AI Pro, and AI Ultra Actually Cap, and When They Reset

Gemini usage limits: the short answer

As of September 2026, Google publishes no prompt count for any Gemini plan. What it publishes is a compute-based limit that refreshes every 5 hours until you hit a weekly cap, and a set of multipliers: the free tier gets "standard limits," Google AI Plus ($4.99) gets 2x, Google AI Pro ($19.99) gets 4x, and Google AI Ultra gets 5x AI Pro at $99.99 or 20x at $199.99. Media generation, Deep Research, the Pro model, and Deep Think burn the allowance fastest, and paid users who hit the wall can keep chatting on Flash-Lite. Everything below comes from Google's Gemini subscriptions page and the Gemini Apps help center, read on September 15, 2026.

If you came here looking for "50 prompts a day," that number stopped existing on May 17, 2026, when Google moved Gemini to compute-based metering. Here's what replaced it. The same question for every other vendor, side by side, is on AI usage limits by plan.

Gemini usage limits by plan

Plan Price Usage limit Context window What's included Window and reset
Free $0 Standard limits 32K tokens 3.6 Flash, "varying access" to 3.1 Pro, image generation, Deep Research, Gemini Live, Canvas, Gems Refreshes every 5 hours until the weekly limit
Google AI Plus $4.99/month 2x Free 128K tokens Everything in Free plus video generation, Daily Brief, 200 Flow credits, 400 GB storage Refreshes every 5 hours until the weekly limit
Google AI Pro $19.99/month 4x Free 1M tokens Everything in Plus, 1,000 Flow credits, Gemini Spark, higher Jules limits, 5 TB storage Refreshes every 5 hours until the weekly limit
Google AI Ultra 5x $99.99/month 5x AI Pro 1M tokens Everything in Pro, Deep Think, 10,000 Flow credits, 20 TB storage Refreshes every 5 hours until the weekly limit
Google AI Ultra 20x $199.99/month 20x AI Pro 1M tokens Everything in Ultra 5x, 25,000 Flow credits, 30 TB storage Refreshes every 5 hours until the weekly limit
Workspace Business Starter / Standard / Plus $7 / $14 / $22 per user/month, annual Not published Not published Gemini in Gmail (Starter); Gmail, Docs, Meet and expanded Gemini app access (Standard and up) Not published
Gemini Enterprise, Business edition From $21/seat/month Not published Not published Gemini Enterprise app, connectors, no-code agents, up to 300 seats Not published

Two things the table can't show. First, the multipliers are relative to a "standard limit" Google never states in absolute terms, so 4x is a real ratio applied to an unknown base. Second, all three models (Flash-Lite, Flash, and Pro) are available on every tier including Free; what the paid tiers buy is more compute against them, not access to them. My verdict on whether the $19.99 tier earns its keep is on is Gemini worth it.

How the Gemini meter works

Google's help center describes the limit in one sentence: "compute-based usage limits that determine how much you can interact with Gemini tools and features. These limits factor in the complexity of your prompt, the models and features you use, and the length of your chat." Three consequences follow.

  • The model you pick sets the burn rate. Flash-Lite is the "efficient workhorse," Flash balances speed and reasoning, and Pro is "our most advanced model" with longer response times. Google says outright that "more advanced models and higher thinking levels consume more of your usage."
  • Chat length compounds. A long thread costs more per turn than a short one, because the whole conversation is context. Starting a new chat for a new topic is the cheapest habit on this page.
  • Features are priced in compute too. Google lists the expensive ones: image, video, and music generation, Deep Research, the Pro model, and extended thinking or Deep Think. Deep Think is Ultra-only and requires the Pro model, so it's the most expensive thing you can ask Gemini to do.

You can see where you stand at gemini.google.com under Settings, then Usage Limits. Gemini also notifies you when you're close, and again when you hit it, with the refresh time.

The 5-hour refresh and the weekly cap

Gemini uses the same two-layer structure OpenAI and Anthropic use, and Google's wording is the clearest of the three: "Your limit refreshes every 5 hours until you reach your weekly limit." So there are two ceilings. The 5-hour one refills on its own during the day. The weekly one sits above it, and once you've spent that, the 5-hour refresh stops helping until the week rolls over. Google doesn't publish how large the weekly limit is on any tier, or exactly when it resets, so I'm not going to invent either.

The practical read: a Pro subscriber who runs three Deep Research reports and a video generation before lunch can hit the 5-hour ceiling, wait, and be fine by mid-afternoon. A subscriber who does that every day will meet the weekly ceiling instead, and that one doesn't refresh at 3pm.

Deep Research, video, and images

None of these carries a published count. Google's feature table shows Deep Research available on every tier including Free, video generation on the paid tiers, and image generation with Nano Banana 2 everywhere, but the only quantitative statement is that they "consume more of your usage." Two tier-specific notes do appear. For users without a Google AI plan, "some features that require more compute resources, like Deep Research, may be unavailable during periods of high demand." And when capacity changes, "limits for users without a Google AI plan may be limited before users with a plan." Free is first in line to lose Deep Research on a busy day.

Video and image creation outside the Gemini app, in Google Flow, run on a separate credit system that Google does quantify: 200 Flow credits on Plus, 1,000 on Pro, 10,000 or 25,000 on Ultra. That's the one place in Google's plan lineup where a hard number exists.

The context window is the published number

Google is precise about the thing it can measure. The free tier's context window is 32K tokens, AI Plus is 128K, and AI Pro and Ultra are 1M, which Google translates as "up to 1,500 pages of text or 30,000 lines of code." Exceeding the window doesn't stop the chat; it means Gemini stops seeing all of what you gave it, and Google warns responses may "miss connections or details throughout the content." For anyone uploading a contract or a codebase, the jump from 32K to 1M is the real reason to pay $19.99, more than the 4x multiplier.

What happens when you hit the limit

  • Subscribers: "you can continue your conversation with Flash-Lite." The chat keeps going on the cheapest model until your limit refreshes.
  • Anyone: upgrade to a higher plan, or wait for the refresh time Gemini shows you.
  • AI credits: the subscriptions page says you "can extend your limits by purchasing AI credits." Google doesn't publish the price per credit on the pages I read.
  • Limits can move. Google says "limits may change without notice, including due to capacity constraints." Plan for the multiplier, not for a fixed quota.

Gemini vs ChatGPT on limits

The structures are close and the disclosure isn't. OpenAI lists everyday text chat as unlimited on every ChatGPT plan and prints estimates for the metered parts (10 to 100 GPT-5.6 Sol messages per 5 hours on Plus, 200 GPT-6 Pro messages a week on Pro $200). Google lists nothing as unlimited and prints multipliers. If you value knowing your number, ChatGPT tells you more; if you value a $4.99 tier and a 1M context window at $19.99, Gemini is the cheaper ladder. The full ChatGPT side is on ChatGPT usage limits, and the head-to-head on everything else is on Gemini vs ChatGPT.

How I'd stay under the Gemini limit

  1. Default to Flash. Reserve Pro for the tasks that need reasoning across a long document, and Deep Think for almost nothing.
  2. New chat, new topic. The single biggest lever, since chat length is one of the three inputs Google names.
  3. Batch Deep Research. Run reports in one block early in a 5-hour window, so the refresh lands before you need the next set.
  4. Watch the weekly, not the 5-hour. Check Settings, then Usage Limits mid-week. The 5-hour ceiling is a pause; the weekly one is a stop.
  5. Business accounts: don't assume the multipliers. Workspace and Gemini Enterprise publish no limit. Test with one seat for a month before rolling it out.
Free Resource Justin McKelvey

Get the Free AI Content Toolkit

The exact system I use to turn one idea into a month of content — atomization framework, voice template, prompt library, weekly system.

Frequently Asked Questions

What are Gemini's usage limits?
As of September 2026, Gemini uses compute-based limits that refresh every 5 hours until you reach a weekly limit, and Google publishes the tiers as multipliers rather than counts: standard limits on the free tier, 2x on Google AI Plus ($4.99), 4x on Google AI Pro ($19.99), and 5x or 20x AI Pro on Google AI Ultra ($99.99 or $199.99). How fast you use the allowance depends on prompt complexity, the model, the features, and how long the chat is.
How many prompts a day do you get with Gemini?
Google doesn't publish a prompts-per-day number for any plan, and since May 17, 2026 the meter isn't a prompt count at all. It's compute: a short question on Flash costs a fraction of what a Deep Research run or a video generation costs. The only quantities Google prints are the multipliers (2x, 4x, 5x, 20x) and the context windows (32K tokens free, 128K on Plus, 1M on Pro and Ultra).
What happens when you hit the Gemini usage limit?
Gemini warns you as you get close, then tells you when the limit refreshes. If you have a Google AI subscription you can keep the conversation going on Flash-Lite, Google's fastest and cheapest model, or buy AI credits to extend the limit. Free users wait for the 5-hour refresh, and during high demand Google says compute-heavy features like Deep Research may be unavailable to them first.
Does Google AI Pro have unlimited Gemini usage?
No. Google AI Pro at $19.99 a month gives you 4x the free tier's limit, a 1M token context window, and access to features like Deep Research and video generation, but the same 5-hour refresh and weekly cap apply. Google says limits may change without notice, including for capacity reasons, and that paid users are limited after free users when that happens.
What uses up the Gemini allowance fastest?
Google lists the expensive features by name: image, video, and music generation, Deep Research, the Pro model, and extended thinking or Deep Think (Ultra only). Its help center also says more advanced models and higher thinking levels consume more of your usage, and that longer chats cost more per turn. Flash-Lite for routine questions and a fresh chat for each new topic are the two cheapest habits.
Is the Gemini limit different on a Workspace business account?
Google's Workspace pricing page lists Gemini in the Business Starter ($7), Standard ($14), and Plus ($22 per user a month, billed annually) tiers, with Standard and above getting expanded access to models and features, but it publishes no usage numbers for any of them. Gemini Enterprise's Business edition starts at $21 a seat a month and also publishes no limit. The multiplier table above applies to personal Google Accounts only.

More on AI for Business

ChatGPT Business vs Plus (2026): Which One a Small Business Should Actually Pay For

ChatGPT Plus vs Business, as of September 2026: Plus is $20 a month for one person and OpenAI may train on your chats unless you opt out; Business is $20 a seat billed annually or $25 monthly, two seats minimum, with no training by default, admin controls, shared projects for groups, and the Company Knowledge plugin. The side-by-side from OpenAI's pages, and three owner scenarios with a verdict each.

6 min

ChatGPT for Customer Service (2026): What Works, What Breaks, and What It Costs a Small Business

ChatGPT for customer service, as of September 2026: it drafts replies, answers policy questions from your own documents, and triages an inbox for $20 to $25 a seat a month on ChatGPT Business. It can't look up an order, issue a refund, or remember a customer unless you connect the system that holds them. The honest capability line, the plan you need, a setup in six steps, the failure modes, and the point where a real agent takes over.

7 min

ChatGPT Pulse (2026): What It Was, Why OpenAI Retired It, and How to Get the Morning Brief for Your Business

ChatGPT Pulse, the daily research cards OpenAI previewed for Pro users in September 2025, was sunset on June 17, 2026 with a 14-day wind-down. What it did, which plans ever had it, why scheduled tasks replaced it, and how an owner rebuilds the morning brief on a Plus or Business seat without handing ChatGPT the wrong things.

7 min

ChatGPT Enterprise Pricing (2026): What OpenAI Publishes, What It Doesn't, and When Business Is the Better Buy

As of September 2026, OpenAI does not publish a ChatGPT Enterprise price: the pricing page says custom pricing and sends you to sales. What it does publish is the shape of the deal (annual contract, seat fees plus credit-based or token-based usage, the per-token rate card), the control set Enterprise adds over ChatGPT Business, and the Business seat prices ($20 to $25 Standard, $100 to $125 Premium) that most companies under 200 people should be comparing against instead.

9 min
Justin McKelvey, Fractional CTO and AI consultant in Austin, TX

Written by

Justin McKelvey

Fractional CTO & AI consultant in Austin, TX. 15 years building software, 50+ products shipped, $53M+ in client revenue generated. I help $1M–$50M founders ship production software and automate operations with AI — without hiring a full-time executive team.

Work with me

If this was useful, here are two ways I can help: