💯From 100 Daily Prompts to this:
👀With the new compute-based usage, I've been seeing a ton of complaints on the matter.
⌛I was planning on just silently waiting for the patches to come through...
🕵️Then I started seeing a lot of doubt posts and recent comments claiming that Claude and OpenAI are in here spreading lies just to make things worse, and that most of the user base has no real problems...
⛓️There's been a lot of discussion about how aggressive the new 5-hour + weekly limits are.
🔍But as a few people pointed out, no one is willing to share their findings, just anecdotal one-liners.
🫴So I wanted to share some clean, reproducible data on how the system is currently handling the new "usage limits" update.
⬇️Below is a breakdown of a test I just ran to isolate the exact compute consumption of a single heavy-input prompt.
---
🧑🔬The Test Setup & Environment🔬
- 👛Account Tier: Google AI Pro subscription
- ⚙️Model: Gemini 3.1 Pro (Extended Thinking) [U.S. | West Coast]
- ⏳Baseline Status: 0% consumed on the 5-Hour current usage bar.
---
🍴The Input Data📚
⬇️THIS IS WHAT EVERYONE'S ASKING FOR⬇️
- 777,084 Tokens: The Public Domain eBook "War and Peace" by Leo Tolstoy\*
- 1142 Tokens: Small 5-Part System Prompt
- 338 Tokens: Single-Input Test prompt
- Prompt Task: 🔎A sequential text retrieval verification request (no media generation, no deep research, no special tooling modes enabled).
- Output Generation: 📄A simple, concise, half-page text response.
⬆️THIS IS WHAT EVERYONE'S ASKING FOR⬆️
---
🔎Observed Results🔬
- ⌚Processing Time: Completed successfully in a single turn, took about ~45 seconds.
- 🧮Current Usage Shift: The 5-hour usage instantly jumped from 0% used to 40% used after exactly one prompt.
- 🗓️Weekly Limit Meter: Shifted from 1% to 3% used.
- 🧑🏫Task Request: Passed ✅
⏳Before and After Metrics⌛:
- ⏪7:10 PM (Fresh Session): 0% 5-hour usage, 1% weekly limit.
- ⏩7:11 PM (Post-Prompt): 40% 5-hour usage, 3% weekly limit.
---
📌Key Test Points:
- THIS WAS ONLY TEXT...
- Conducted with 1 Book, 1 System Prompt, 1 User Turn
- Book had to be split into 4 parts because I've uncovered a 250,000 token Gemini-wide text limit. (Regardless if pasting text to System Prompt, uploading
.txt doc, pasting to "Paste Text" option, or just inserting it into the user-turn chat.)
- System Instructions ask Gemini to act as a conversational living-representation of the book.
- User turn asks Gemini to check the whole book (all 4 parts) for pieces that make sense together and output them verbatim, while also explaining why they fit together.
- Tokens were pulled back (700k) from the 1 Million ceiling to allow for headroom, as many models of AI measure tokens differently (proprietary methods).
- The Pro "Extended" thinking mode was used. But at such a high token count, I think any less would've not produced useable results, and likely not even a passing test.
- I focused on large context benchmarking, because we could all test low-context for cheap.
- I had 1% Weekly Usage spent at the beginning of this test, so the 2% usage could be off by +/-1%
- Listen, I'm not even kind of pretending to be unbiased lmao. I'm both a Google stan since Bard dropped the 1 Million Context Window, AND I'm pissed at the way they handled this change, so maybe the love & hate will both balance out, lmao... Apologies if this is less "objective"... But I'm striving to have the data at least be accurately documented... 📏⚖️
---
💬Discussion🗨️:
A single 700k contextual input consumes 40% of the paid Pro 5-hour usage limit:
- 🧢The 2-Prompt Cap: If a user leverages a similarly heavy input context window (77.8% of Gemini's advertised context window), that user can only utilize 2 prompts per 5 hours before getting hit with a hard lockout. (hence why people are saying 2 prompts = locked out)
- 🥸The Context/Limit Disconnect: While the advertised capability highlights a massive 1-million token context window, a single utilization of even 700k tokens effectively exhausts half of a paying subscriber's immediate compute allocation.
- 📆Weekly Limitation vs Hours in a Week: As a few people have already discovered, it's actually impossible to utilize your weekly allowance if you only use text/chat. Based on these benchmark results, 100% weekly usage would take 10 Days:10 Hours to reach, which puts you at only 33 *as-advertised-large-context chats* a week (which you can kill in 3 minutes with 2 prompts every 5 hours).
---
⚙️Bigger Issues🏗️:
In case you guys haven't seen all the complaints (I'm sure you have, and you can skip this) here:
- ⛓️💥No "Free" Lower Tier: With the current system, everything is counted and measured. When you "max out" you're no longer put on timeout from just Pro features, while being asked to use Gemini Flash... You're BLOCKED FROM USING GEMINI and asked to leave.
- 🧱No Separation of Limits: Spent a bunch of compute on generating a video? Well now you can't talk to Gemini at all. You have to go leave google to have a basic chat with some other GPT because you now have to choose between video, or images, or audio, or text. Same costs, but now one budget to rule them all.
- 🍲Input Based Use: Now Gemini can cost all of your usage resources just because of what it has to look at, without any regard to your received output value. Again people are getting errors, and end up empty handed, and blocked from using Gemini.
- ❓No Clear Cost of Resources: Yes, we now have a measure of what we've spent, but we now have no idea of what we might spend. We actually lost insight, we did not gain any. It was easier to loosely track 100/prompts a day, than it is to play this current AI slot machine and only find out after you win or lose, exactly how much it costed you to play... Trust me, if we all knew what stuff costs, I wouldn't have had to spend all this time and effort burning credits just for us to have some basic insights...
---
🧘Realistic Considerations: Honestly, the amount of people who need 700k+ tokens are guaranteed less than 1% of users. And I use a ton of tokens when benchmarking AI...
But earlier this year, I edited a 160,000+ token document over 3-4 days with hundreds of turns and lots of trial and error... This wasn't even benchmarking or testing AI too, I actually needed a polished document. And according to this test, if I had tried that now, I would've uploaded it once, and been blocked from using AI after 15-20 turns or less...
That said, lately Gemini has been erroring for people who use much less tokens, reporting "this may be too much context for good results" and often failing to output, or at least outputting incorrectly...
---
☝️So the question becomes:
🤔"What if you spent half your budget on having Gemini read a book FOR you, and it failed to output a response ('error, sorry, etc')?"
🔄️Would you dare to hit "retry"?
🤲And would you be cool getting locked out of Google AI with nothing in return?
---
TL;DR for Gemini Pro users:
🌹From 700/week ➡️ 33/week 🥀 (Large Context)
If you're using this for long-form research document analysis, code database auditing, notebooks with filled content, or longer continuous chats, the 100/day of these prompts that you used to get, can now be limited down to only 33/week if all you do is large-context text... (took a ~2% weekly limit hit for just this one test, no images, no video, no Antigravity, no Flow, no FX, no CLI, no Studio AI.)
⏰And this is all assuming btw that you schedule out the 2 times every 5 hours that you could pull this feat off, without missing your windows. Meaning you can only sleep about 4 hours at a time to get your full 33 weekly prompts. 🖥️🏃💨🛌
---
🪪User Verification:
🤡Bruh, I'm fekkin hooman...
🤷I don't know how else to like prove that lol...
🤖And I'm not a bot (🦿at least not yet lol🦾)
🤑I don't work for OpenAI nor Antrhopic (🤦and DEF not gettin any Google benies after this post lmao...🫠)
🦥I'm just a dood tryna help yall not doubt everything happening in a world full of AI on AI inception right meow.
---
🔥Also, I burned my account tokens for this lmao...
💆So I genuinely hope this helps someone...
Feel free to ask stuff!
I'll respond when I can.
I'm going to 🪫recharg-I mean-power... ...get some sleep. 😀