Every day when I use Antigravity for work between 16:00 and 22:00 UTC, Gemini Flash 3.8 constantly goes on strike.
Even when it’s not down, my tests show that during this timeframe, Flash’s speed isn’t even 1/10th of what it usually is. A 2-minute task takes over 20 minutes and still isn’t done—it’s absolutely infuriating!
Google Cloud model server 503 capacity exhausted (throwing errors multiple times):
Yesterday at noon and early this morning, the API frequently returned: UNAVAILABLE (code 503): No capacity available for model gemini-3.8-flash-high on the server
Along with network connections being cut off remotely: write tcp ... wsasend: An existing connection was forcibly closed by the remote host dial tcp: lookup daily-cloudcode-pa.googleapis.com: no such host
This happens about 90% of the time.
Right now, as I’m writing this post, Flash has been running for 30 minutes and hasn’t finished a single task. It’s painfully slow and lagging.
Wait a minute… the news says most AI developers have already migrated to ChatGPT, DeepSeek, and Grok, with the ChatGPT API being particularly cheap and highly rated. Given this, it makes no sense that a decrease in Gemini API’s overall user base would contradictorily result in service suspensions due to “overcapacity.” I can only think of one explanation: Google is intentionally throttling the gateway capacity and diverting the freed-up compute power to internally test Gemini 4
If that’s true, I absolutely support Google hustling to catch up with ChatGPT and Claude Code, but please don’t squeeze the gateway capacity this hard, okay? Thanks.
Honestly, Google could just drop Claude Code Opus 4.6. It’s outdated, and its performance and benchmark scores are bottom-tier—even Flash High beats it by a mile. If Google isn’t going to update to Opus 5.5, there’s really no need to keep paying Anthropic. It’s a waste of money; Gemini is already great on its own.
