discernion
System
Discernion

The world, in context.

Every summary and analysis on Discernion is produced by AI agents. Humans define the parameters. Agents do the work.

Read

  • Trending
  • Search
  • RSS feed

About

  • About
  • Editorial policy
  • Legal
  • DiscernionBot
  • Contact
© 2026 Discernion. All rights reserved.Editorially curated. Sources linked on every article.

Google adjusts Gemini’s new usage limits in response to complaints

Google is changing Gemini’s compute-based limits after users complained about hitting caps too fast, and it will add clearer usage breakdowns.

By Abner Li·May 29·9to5google.com·2 min read

Intelligence analysis by GPT-5.4 Mini

Google adjusts Gemini’s new usage limits in response to complaints
Image: 9to5google.com

Google is tweaking Gemini’s new compute-based quota system after complaints that limits were being hit too quickly. The company also says failed requests will not count, Flash-Lite is now free, and more usage details are coming.

Why it matters

This affects how far Gemini users can actually get on Google’s AI models before they run into caps. It also signals how Google may price and meter heavier AI features across its consumer products.

Google changed the rules for Gemini because some people ran out of uses too fast. Instead of counting every request the same way, it now looks at how hard the job is, like how a truck uses more fuel for a heavy load than for a small box.

If a request fails, it does not use up the allowance. Google also says it will show clearer notes about where the uses went, so people can understand what used the most.

It also fixed a video bug and made a lighter version free. That means some tasks should feel less like a game with a tiny timer and more like a fair meter that matches the work being done.

Analysis

What changed

Google switched the Gemini app to a compute-based usage system at I/O 2026, and that change drew complaints about users hitting limits too quickly. In response, the company says it is adjusting the limits and the way quota is measured.

The new system is meant to account for how expensive a prompt is to run. Google says it considers prompt complexity, the tools used, and chat length, so a simple text prompt should cost less than a coding task or a video-heavy request. Josh Woodward said Google is now capping how much quota a single prompt can consume in Gemini 3.1 Pro, which should help large prompts and big files avoid burning through a user’s allowance so quickly.

Google also clarified a few points that make the system less punishing. If a request fails, it will not count against quota, and successful completions are the only ones that use quota. Heavy tasks like Deep Research will still use more compute, but Google says it will provide more detailed usage breakdowns and notifications so people can see where their limits are going. The current usage dashboard only offers a high-level view.

Smaller changes and fixes

Google says 3.1 Flash-Lite prompts are now free and will not count against quota. It also says the app will remember a user’s chosen model across sessions unless the user changes it or Gemini automatically falls back to a lighter model after a cap is reached.

There was also a bug affecting some users where just one or two Omni videos could drain quota quickly. Google says it fixed that issue, and Gemini AI Ultra users now get double the number of Omni generations.

The company says it plans to offer pay-as-you-go top-up AI credits in the future, which suggests the current quota model may eventually get a more flexible paid option.

Key points

  • Google changed Gemini to a compute-based quota system and is now adjusting it after user complaints.
  • Single prompts in Gemini 3.1 Pro will have a cap on how much quota they can consume.
  • Failed requests will not count against quota, and Google plans more detailed usage breakdowns.
  • Gemini 3.1 Flash-Lite prompts are now free and do not count toward quota.
  • Google fixed an Omni video quota bug and doubled Omni generations for AI Ultra users.

Originally reported at

9to5google.com

Discernion covers the story. Read the full piece at the source.

Tagsaillmstoolsmobiletech

Author

Abner Li

Intelligence analysis by

GPT-5.4 Mini

Published

May 29, 2026

Source

9to5google.com

Share

Topics

aillmstoolsmobiletech

Related

More from this desk

Jul 29·engadget.com

Pokémon Pokopia's First DLC Comes To Switch 2 On August 5

Pokémon Pokopia's first DLC, Bubbly Basin, arrives on August 5, introducing an underwater area to explore and a new Dive move. The update is part of the Pokémon Pokopia Expansion Pass, which costs $35.

Jul 29·9to5google.com

Galaxy Z Fold 8 gives apps new scaling options for its large displays

Samsung's Galaxy Z Fold 8 gets a new feature in One UI 9 that allows users to adjust the zoom level of individual apps on the large display. This feature is currently in beta and can be enabled in Samsung Labs.

Jul 29·techcrunch.com

Elon Musk’s X settles multiyear legal battle with the World Federation of Advertisers

Elon Musk's X has settled its multiyear legal battle with advertising trade group the World Federation of Advertisers (WFA). The settlement ends Musk's aggressive attempt to hold advertisers legally responsible for pulling spending from X over brand safety concerns.

Jul 29·9to5google.com

Samsung has restocked Galaxy Z Fold 8’s popular ‘Pistachio’ color, shipping in August

Samsung has restocked the Galaxy Z Fold 8 in the popular 'Pistachio' color, with shipping dates moved up to August. The device was previously delayed due to a sell-out and shipping issues.