gen‑ai.news
← Back
Video

Google fixes several bugs in Gemini usage limits that burned through quotas too fast

Google fixes several bugs in Gemini usage limits that burned through quotas too fast

Google has addressed multiple bugs in the Gemini app that were incorrectly depleting users' monthly usage quotas far faster than intended. The most significant issue involved the Imagen-powered Omni video feature, where generating just one or two clips could exhaust an entire quota allotment - a clear sign that the cost accounting behind the scenes was miscalculated rather than reflecting actual resource consumption.

As part of the fix, Google has doubled the number of video generations available to Gemini Ultra subscribers, which suggests the original quota was already set too conservatively even before the bug compounded the problem. The company has also stopped billing users for requests that fail outright, a practice that had drawn frustration since users were losing quota credit for content they never actually received.

These kinds of quota bugs matter more than they might initially appear. As generative video becomes a more routine part of AI-assisted workflows, predictable and accurate usage tracking is necessary for users to plan their work and for Google to maintain trust in its subscription model. When quotas behave erratically, users have little basis for understanding what they can actually do within a given billing period.

Google says it also plans to introduce greater transparency around usage limits more broadly, though specifics on what that will look like have not been detailed yet. For now, the immediate fixes should give Ultra subscribers a noticeably more generous and reliable experience with Omni video generation going forward.

Enjoy this story? Get the next one in your inbox.

Twice a week: the most important stories in generative image and video AI, distilled into a 2-minute read.

Free. Unsubscribe any time. No spam, ever.

Your next read

No image
Video

Runway’s WorldPrompt and the Engineering of Real-Time Worlds

Runway's WorldPrompt system powers its Gen World Models 2 (GWM 2), enabling real-time generation of video and audio through persistent context and timed actions. Rather than producing discrete clips, the model maintains a continuous understanding of an environment as it unfolds. The approach marks a notable shift in how world models can be steered interactively.

Gemini 3.8 Live with Live Avatar gives Google’s AI a face
Video

Gemini 3.8 Live with Live Avatar gives Google’s AI a face

Google has updated Gemini Live with an animated avatar that lip-syncs and displays facial expressions in real time during conversations. Called Live Avatar, the feature is currently limited to Gemini Enterprise customers and supports 97 languages without degrading video quality. It marks Google's latest step toward giving its AI assistant a visible, expressive presence.

No image
Video

Introducing Gemini 3.8 Live with Live Avatar

Google DeepMind has introduced Gemini 3.8 Live, an updated multimodal model paired with a new Live Avatar feature that generates an animated, talking on-screen presence during real-time conversations. The combination allows users to interact with a responsive visual agent rather than a purely voice-based interface. The release marks another step in Google's effort to make AI interactions feel more immediate and embodied.