Does Gemini have a limit? Yes, and since May 17, 2026 it stopped being a number you can count. Google retired daily prompt caps and replaced them with compute-based usage limits that refresh every five hours inside a weekly cap. The same twenty questions might cost you nothing on Monday and half your allowance on Tuesday, depending on which model answered them and how long the chat had been running.
That change made a simple question surprisingly hard to answer, because Google now publishes multipliers rather than numbers. This guide covers what the allowance actually measures, what each plan gets, when it resets, how to see where you stand, what drains it fastest, and what happens at the wall. It also corrects one claim that has spread widely and is wrong.
The Key Takeaways
- Yes, there is a limit, but not a prompt count. Gemini meters compute, weighing prompt complexity, the model and features you use, the size and number of files you upload, and chat length.
- It refreshes every five hours, inside a weekly cap. Two clocks run at once, and the weekly one is the one that stops you.
- Google publishes multipliers, not counts: standard on Free, 2x on AI Plus at $4.99, 4x on AI Pro at $19.99, and 5x or 20x on AI Ultra.
- Ultra's multiplier is measured against AI Pro, not against Free. The $199.99 tier is 20x Pro, which works out to roughly 80x the free allowance.
- Buying extra Gemini usage is not settled. Google’s AI credits documentation covers Google Antigravity and Google Flow only, while a footnote on its own subscriptions page claims you can top up the Gemini app.
Does Gemini Have a Limit? The Short Answer
Yes. Every Gemini plan, including the free one, runs against a usage limit, and Google is explicit that it reserves the right to cap how many prompts and conversations you get inside a given window.
What changed in May 2026 is the unit. Instead of counting prompts, Gemini now measures compute, which Google describes as factoring in the complexity of your prompt, the models and features you use, and the length of your chat. Google's work and school documentation adds one more driver that the consumer page leaves out: the size and number of files you upload. Your allowance refreshes every five hours until you reach a weekly ceiling.
The practical consequence is that no honest answer to the how-many-prompts question exists any more. A short text question to the lightest model and a video generation request both draw on the same pool at wildly different rates.
What Replaced Gemini's Daily Prompt Limits
The reasoning Google gave in its Google AI subscriptions announcement at I/O 2026 was one line long. A simple text prompt draws far less compute than a complex video or coding prompt, so charging both the same made no sense.
The rollout came in two waves. Users over 18 moved on May 17, 2026. Users under 18 moved on July 24, 2026, a second date that almost no coverage mentions.
The Numbers Google Used to Publish
Before the switch, Google was far more forthcoming. Its help page as archived on 1 April 2026 carried a full table of daily counts, and those numbers are worth keeping on the record because the current page has nothing like them.
| Model (pre-May 2026) | Free | AI Plus | AI Pro | AI Ultra |
|---|---|---|---|---|
| Pro | Basic access, no number | 30 prompts/day | 100 prompts/day | 500 prompts/day |
| Thinking | Basic access, no number | 90 prompts/day | 300 prompts/day | 1500 prompts/day |
Two things stand out. Google quantified the paid tiers but never the free one, which sat at an unnumbered "Basic access" then just as it sits at unnumbered "standard limits" now. And the figure repeated across the web, that free users once got five prompts a day, is a misreading of that same table. The only fives in the free column were 5 screen automation requests a day and 5 Deep Research reports a month. Neither is a prompt allowance.
All of it went when compute metering arrived. What Google publishes for consumers today is the multiplier table below, a context window per plan, and nothing else.
Gemini Usage Limits by Plan in 2026
Here is the whole published picture for consumers, and the thing to notice is what is missing from it. Every allowance is expressed as a multiple of the free tier's "standard limits", a quantity Google has never put a number on. Prices are the monthly US rates from Google's subscriptions page.
| Plan | Price | Usage allowance | Context window |
|---|---|---|---|
| Free | $0 | Standard limits | 32k tokens |
| Google AI Plus | $4.99/mo | 2x Free | 128k tokens |
| Google AI Pro | $19.99/mo | 4x Free | 1 million tokens |
| Google AI Ultra | $99.99/mo | 5x AI Pro | 1 million tokens |
| Google AI Ultra | $199.99/mo | 20x AI Pro | 1 million tokens |
The context window is the one hard number in the whole system, and it is worth reading as a limit in its own right. At 1 million tokens Google reckons Gemini handles about 1,500 pages of text or 30,000 lines of code in a single sitting. At 32k on the free plan, long documents get truncated in ways that quietly degrade answers rather than throwing an error. Our guide to AI file upload limits across the major chatbots covers where that bites hardest.
Why Ultra's Multiplier Is Bigger Than It Looks
This is the detail most summaries flatten. Google measures AI Ultra against AI Pro, not against the free tier. Google’s own limits documentation says Ultra is "5x or 20x higher than AI Pro limits", and its subscriptions page repeats the wording price by price.
Since AI Pro is itself 4x the free allowance, the $199.99 Ultra tier lands at roughly 80x free by simple arithmetic, and the $99.99 tier at about 20x. Google does not state those compounded figures anywhere, so treat them as our calculation rather than a published number. They do explain why the jump from Pro to Ultra costs so much more than the jump from Free to Pro.
What the Free Plan Actually Reaches
Model access is not the dividing line people expect. Google's plan table marks Flash-Lite, Flash and the Pro model as available on every tier, free included. What separates the plans is how much of the expensive models you get through before the meter runs out, plus a handful of gated features like video generation and Gemini Spark.
We cover what the $0 tier includes in full in our breakdown of the free Gemini plan, and the complete price ladder across consumer, API and Workspace in the Gemini pricing guide.
When Do Gemini Limits Reset?
Two clocks run at the same time, and confusing them is the most common reason people think their limit is broken.
The Five Hour Refresh
Your allowance refreshes every five hours. This is the pause, not the stop. Hit it at ten in the morning and you are working again by mid-afternoon, which for most people is an inconvenience rather than a blocker.
The Weekly Cap
Sitting above that is a weekly ceiling, and Google is clear that the five-hour refresh only continues until you reach it. That makes the weekly number the one that decides whether a plan fits you. A heavy Tuesday and Wednesday can leave you rationing on Friday, with no five-hour reset able to rescue you.
Image Limits Reset Daily Instead
Image generation and editing do not follow the five-hour rhythm. Google has said in both its old consumer table and its current work and school documentation that these limits change frequently and reset daily, because demand for them runs so high. If your image quota is gone, the wait is until tomorrow, not until this afternoon. For what the underlying image models cost elsewhere, see our Nano Banana pricing breakdown.
How to Check Your Gemini Usage Limits
Google exposes a usage view, which is more than most of its rivals offer, and almost no guide to this topic mentions it. It takes two clicks.
- Go to gemini.google.com.
- At the bottom left, select Settings, then Usage Limits.
You also get warned twice without looking. Gemini notifies you when you are close to your limit, and again when you hit it, and the second notification tells you when the allowance refreshes.
Your Usage Pools Across Every Surface
One detail catches people out. Gemini on the web, the mobile app and Gemini in Chrome all draw on a single allowance. Google spells this out in its work and school documentation rather than on the consumer page. Use Gemini in Chrome, then switch to the phone on the same model, and both count against the same pool. There is no per-device budget to spread work across.
What Drains Your Gemini Allowance Fastest
Google names the expensive things directly, which makes rationing straightforward once you know the order.
| What you do | Why it costs more |
|---|---|
| Media generation | Images, video and music are the heaviest single requests Gemini serves. |
| Deep Research | One run reads across many sources before it writes anything. |
| The Pro model | The most capable of the three model tiers, and the most expensive per answer. |
| Extended thinking and Deep Think | The model reasons for longer before it responds. |
| Long chats and large uploads | Chat length, file size and file count all feed the calculation. |
Google's own guidance is blunt about the pattern: more advanced models and higher thinking levels consume more of your usage. Deep Think is the extreme case, restricted to AI Ultra and dependent on the Pro model underneath it. If you run research workloads regularly, our Deep Research workflow guide is built around getting more out of each expensive run.
What Does Not Cost You Anything
Two rules work in your favour, and both came out of the adjustments Google made after users complained about burning through the new limits.
Failed requests are free. Gemini lead Josh Woodward, as quoted by 9to5Google, said that a failed request is not charged to you. System mistakes are Google's rather than yours, and quota is consumed only by successful completions. Google also capped how much quota any single prompt can consume on the Pro model, so one enormous file no longer wipes out an afternoon.
One more piece of behaviour is worth knowing: Google says your model choice sticks across sessions, changing only if you change it yourself or if a cap forces an automatic fallback to something lighter.
How to Make the Allowance Last
Three habits follow directly from how the meter works. Watch the weekly number rather than the five-hour one, since that is the clock with no way back; a mid-week look at the usage screen tells you whether Friday is at risk. Keep routine questions on the lighter models and save the Pro model and extended thinking for work that actually needs the reasoning.
The third one is the least obvious. Long chats cost more as they grow, because the whole conversation feeds the calculation, so starting a fresh chat for an unrelated task is cheaper than continuing a sprawling one. The same logic applies to uploads: fewer and smaller files, fewer tokens weighed against you.
What Happens When You Hit the Limit
You do not get locked out of Gemini. You get demoted.
If you hold a Google AI subscription and run out, Google lets you carry on in the same conversation using Flash-Lite. Your model choice is remembered across sessions, so the demotion is temporary: once the heavier model’s limit refreshes you are back on it. Otherwise the documented options are the plain ones: upgrade, or wait for the refresh.
No, You Cannot Buy Extra Gemini Usage Yet
This is the claim worth correcting, because several competing guides tell readers they can top up their Gemini limits with AI credits. As of September 2026 that is still wrong.
Google's I/O announcement says AI Pro and Ultra subscribers can buy pay-as-you-go top-up AI credits for Google Antigravity and Google Flow, and lists the Gemini app as coming soon. Google's help pages on AI credits still name only those two products. The live limits page backs that up by omission: when it sets out what to do at a cap, it offers two routes, upgrading or waiting.
There is a genuine contradiction to flag, though, because Google does not speak with one voice here. Footnote 2 on the Gemini subscriptions page, under the heading "Usage access limits in the Gemini app", states that you can extend your limits by purchasing AI credits. The Help Center page that footnote links to does not mention credits at all.
So one Google page sells a top-up that its own documentation does not describe. Until that resolves, plan as though the levers at a cap are a bigger plan, a lighter model, or the clock. Check what you are actually buying before paying for credits to lift a Gemini limit.
The Free Tier Gets Squeezed First
There is an asymmetry in the fine print that free users should know about. Google warns that when Gemini is busy, compute-hungry features like Deep Research might go unavailable for people without a paid plan, and that if capacity tightens, free users risk being restricted before paying ones. Both are worded as possibilities rather than promises, which is itself the point: the free allowance is the one Google reserves the right to squeeze.
So the free allowance is not only smaller, it is softer. It shrinks exactly when everyone wants it, which is the hidden cost of the $0 tier and the strongest practical argument for AI Plus at $4.99 if Gemini is part of your working day.
Does Gemini Have a Limit on Work and School Accounts?
Yes, and here Google does something it refuses to do for consumers: it publishes actual numbers. If you use Gemini through a Workspace account, you get counts rather than multipliers.
| Workspace tier | Pro model | Thinking model | Context |
|---|---|---|---|
| Business Starter, Essentials, Frontline | Basic access, varies | Basic access, varies | 32k |
| Business Standard and Plus, Enterprise Standard and Plus | Up to 25 prompts per 4 hours | Up to 300 prompts/day | 1 million |
| AI Expanded Access | Up to 200 prompts/day | Up to 600 prompts/day | 1 million |
| AI Ultra Access | Up to 500 prompts/day | Up to 1500 prompts/day | 1 million |
Note the shape of that Business Standard row. Twenty-five Pro prompts per four hours is a tighter leash on the best model than many paying consumers expect, even though the context window jumps to a million tokens. Deep Think, where offered, runs to 10 prompts a day. Google also notes separately that using Gemini inside Workspace apps like Docs and Gmail carries its own limits, distinct from these.
How Gemini Limits Compare With ChatGPT and Claude
All three major assistants have converged on the same basic design: a rolling short window stacked underneath a longer one, rather than a daily reset. What separates them is disclosure.
Gemini is the least specific of the three for consumers. It gives you a refresh cadence, a set of multipliers and a context window, and nothing you could budget against in advance. The trade is that it also gives you a usage screen and two warnings, which is more visibility than the multiplier-only documentation suggests. For the equivalent mechanics elsewhere, see how ChatGPT limits reset and how to check them and Claude's session and weekly caps explained. We have deliberately not restated their numbers here, since each vendor moves them on its own schedule.
One Way Around the Single Vendor Ceiling
If you keep hitting a wall on one assistant, the fix is not always a bigger plan on that assistant. Running out on Gemini at 3pm matters far less when the same work moves to another model instead of waiting for a refresh.
That is the case for a multi-model app rather than a stack of subscriptions. Fello AI puts Gemini alongside GPT, Claude, Perplexity, Grok, DeepSeek and others on the Mac for $9.99 a month, so a cap on one model is a switch rather than a stop. Set against $19.99 for AI Pro on its own, it is a different way to spend the same budget, and worth weighing before you upgrade. If you are already paying for several assistants, our AI subscription audit is the place to start.
The Verdict
Gemini's limits are real, deliberately unquantified, and more generous in practice than the silence around them suggests. The free tier reaches every model Google ships, which is unusually open, but it is also the first thing throttled when demand spikes, so treat it as a casual plan rather than a dependable one.
For daily work, AI Pro at $19.99 is the tier that stops you thinking about the meter, and the $4.99 Plus plan is the better buy if your use is regular but light. Ultra earns its price only for sustained heavy generation, and its real multiplier against free is far larger than the headline 20x implies.
The one thing to fix in your mental model: you probably cannot buy your way out of a Gemini cap, whatever the subscriptions page footnote implies. Until Google’s credits documentation names the Gemini app, the reliable levers are a bigger plan, a lighter model, or the clock.
Frequently Asked Questions
Does Gemini still have a daily limit?
No, not for general use. Since May 17, 2026 Gemini has run on compute-based limits that refresh every five hours inside a weekly cap. Image generation is the exception: those limits still reset daily.
How many prompts do I get on the free Gemini plan?
Google publishes no number for the free plan and never has. Even before May 2026, when its help page did list daily counts for the paid tiers, the free column read "Basic access" with no figure. The widely repeated claim that free users got five prompts a day is a misreading of that old table, where the fives were screen automation requests and Deep Research reports. Today only work and school accounts get published prompt counts.
Can I buy more Gemini usage when I hit the limit?
Google’s own pages disagree. Its AI credits documentation covers Google Antigravity and Google Flow, and the Gemini app has been listed as coming soon since I/O 2026, yet a footnote on the subscriptions page says you can extend Gemini app limits by buying credits. Assume the documented position until Google aligns them: upgrade, switch to a lighter model, or wait for the refresh.
Do failed Gemini responses count against my limit?
No. Google has said that quota is consumed only by successful completions, and that a failed request is not charged to you. Google also caps how much of your allowance a single prompt is able to consume on the Pro model.
Is AI Ultra really 20x the free plan?
No, it is more. Google measures Ultra against AI Pro, so the $199.99 tier is 20x Pro. Since Pro is 4x free, that works out at roughly 80x the free allowance. The compounded figure is arithmetic on Google's two published multipliers rather than a number Google states.