惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

aimingoo的专栏
aimingoo的专栏
宝玉的分享
宝玉的分享
J
Java Code Geeks
Martin Fowler
Martin Fowler
博客园 - Franky
I
InfoQ
Stack Overflow Blog
Stack Overflow Blog
Blog — PlanetScale
Blog — PlanetScale
S
SegmentFault 最新的问题
B
Blog
The Cloudflare Blog
F
Fortinet All Blogs
量子位
腾讯CDC
博客园 - 司徒正美
D
Docker
大猫的无限游戏
大猫的无限游戏
Microsoft Azure Blog
Microsoft Azure Blog
T
The Blog of Author Tim Ferriss
V
Visual Studio Blog
IT之家
IT之家
Last Week in AI
Last Week in AI
D
DataBreaches.Net
小众软件
小众软件

Android Authority

Still thinking about buying the Trump Mobile phone? Here are 5 reasons why you shouldn’t I know YouTube Music is flawed, yet I prefer it over Spotify Survey reveals 50% of users don’t like the new Google Health app It’s time for Samsung’s S Pen to evolve or die The Motorola Moto G Stylus (2026) is a sequel we didn’t need NotebookLM is quickly becoming the podcast app I didn’t know I needed Samsung’s next Galaxy Watch update could finally make your health data useful Google’s Gemini Spark is ready to run your digital errands while your phone is off Telegram’s finally getting an official Wear OS app again Nintendo is back on mobile, and it wants to turn your selfies into minigames Google Drive’s big document scanner overhaul is finally here — don’t overlook its power Spotify will finally give you real profile tools to make music listening more social Acer’s new gaming handheld might dodge the worst of tech inflation Meta is cooking up a new line of smart glasses, and they may not be Ray-Bans ChatGPT is retiring this beloved legacy model in June Is Microsoft Copilot not working? Here’s what’s going on (Update: Back up) Samsung Gallery starts quietly ending OneDrive support ahead of schedule Here’s a first look at custom wallpapers in Google Messages Rivian is pretty sure customers want AI, not Android Auto Leaked iPhone 18 Pro dummy units may have just shown the next Android phone color trend A company spent $500 million in one month after forgetting to set AI usage limits Now even MediaTek’s cheap chips are embarrassing the Tensor G5 in one major area Pixel 10 Pro XL user says Google returned their phone worse than dead The best robot pool cleaners of 2026: Top picks for all budgets and pool sizes Claude Opus 4.8 is more honest, less deceptive, and considerably cheaper Roborock’s Qrevo Curv 2 Flow is ready to mop up the competition — and your filthy floors Google is making it easier to share Gemini chats, media, and more with your team One UI 9 borrows one of the iPhone’s most useful call features This is the biggest mistake Oura is making with the Oura Ring 5 This Verizon user owed $400, but the carrier made an unexpected move
Google may have fixed the issue that was exhausting your ...
Shimul Sood · 2026-05-29 · via Android Authority
The old Gemini app interface.

Brady Snyder / Android Authority

TL;DR

  • Google is fixing major quota complaints in Gemini by addressing bugs and making usage limits more predictable.
  • The company is also changing how heavy usage is counted, while failed requests and Flash-Lite prompts won’t count towards limits at all.
  • To improve transparency, Google is adding better breakdowns for deep research usage and making model selection persistent across sessions.

We recently reported that Google had quietly tightened parts of its AI Pro plan, and users did not take long to notice. People instantly started reporting that their limits were being hit much faster than expected, sometimes within just a few prompts. Google later increased quotas for Antigravity users to calm things down, but that only addressed part of the frustration.

Now, Josh Woodward, Vice President at Google, has responded more directly in a post on X, acknowledging that users were encountering limits sooner than they should. He said the company is now rolling out several fixes designed to make usage more predictable, reduce confusion, and ensure quotas feel more consistent across different types of tasks.

Josh Woodward on Gemini usage limits

One of the biggest fixes involves a bug tied to Omni video generation. In some cases, users were finding that just one or two video prompts were eating up a large portion of their quota. For example, someone experimenting with short clips or testing different styles could suddenly see their allowance drop far more than expected after only a couple of attempts. Google says this issue has now been fixed, and it is also increasing allowances for heavier users. Ultra subscribers, for instance, are getting double the number of Omni video generations starting immediately.

Another area that caused complaints was Google’s Complex 3.1 Pro prompts. These are long, detailed instructions, often accompanied by large file uploads or multi-step reasoning tasks. These prompts were also consuming quotas in a way that felt too aggressive. Google is now changing this by introducing caps per prompt. Instead of one very heavy request potentially draining a large chunk of your usage, the system will now limit how much a single prompt can consume. The idea is to prevent extreme outliers where one task wipes out too much of your monthly allowance.

Josh Woodward on Gemini usage limits

There is also a change that users will likely appreciate in everyday use. Woodward noted that about 1 in 10 requests can fail due to system errors. Earlier, even failed attempts could still count against your quota, which understandably felt unfair. That is now being corrected. If a request fails, it will not be charged against your usage. So if Gemini glitches out while generating a response, that attempt no longer eats into your limit.

Josh Woodward on Flash Lite prompts on X

A notable update is that Flash-Lite prompts will no longer count against quota at all. This effectively turns Flash-Lite into a free layer for lighter tasks. It also subtly encourages users to rely on lighter models when they do not need full reasoning power, which should help stretch the limits of higher tiers further.

Google is also working on more detailed breakdowns and notifications for Deep Research usage. These are the more compute-heavy tasks where Gemini processes large inputs or runs multi-step analysis. Many users currently have little visibility into why their quotas drop faster on some days than others. The goal is to make that much clearer, so users can actually see which types of tasks are expensive and which are not.

Josh Woodward about Deep Research on X

Finally, there is a useful improvement in how model selection works. Once you choose a specific model inside Gemini, the app will remember it across sessions. So if you prefer a particular writing or research setup, you won’t need to select it every time you open the app. The only exception is when you hit a usage cap, in which case the system may automatically switch to a lighter model to keep things running.

These changes definitely feel like Google trying to smooth out a system that had become inconsistent for many users. The limits are still there, but the company is clearly trying to make them feel more logical. Whether that fully fixes the frustration remains to be seen, but at least the direction now feels more user-friendly than opaque.

Thank you for being part of our community. Read our Comment Policy before posting.