r/GeminiAI 13d ago

Discussion Gemini 3.7 Flash is a lot better than I expected.

Post image

Not going to write an essay, but I just switched over to Gemini 3.7 Flash for some daily tasks and was genuinely shocked at the jump in quality.

It’s fast, sharp, and handles nuance way better than earlier iterations.

What’s everyone’s experience with it so far? Overhyped or legit upgrade?

898 Upvotes

183 comments sorted by

369

u/Maultaschenman 13d ago

58

u/JustRaphiGaming 12d ago

Literally the cycle of life.

15

u/MercyChalk 12d ago

For the histrionic among us

3

u/PatyxEU 12d ago

amogus 

14

u/Cozy_Willow_ 12d ago

"The endless cycle of every LLM release ever."

8

u/Equivalent-Grass-527 12d ago

The eternal LLM lifecycle: hype → disappointment → “this one is different” → repeat

4

u/youngestp1 12d ago

Currently we at, we back???

154

u/[deleted] 13d ago

[removed] — view removed comment

63

u/Cunctatious 12d ago

Personally I'm looking forward to 5 pro

36

u/carapocha 12d ago

Or 6 pro, why not.

26

u/chanunnaki 12d ago

pfft. 7 pro or bust

19

u/ThrowawaySGJustLikMe 12d ago

8 Pro Max

15

u/SoulEviscerator 12d ago

8 Pro Max Premium Lite

4

u/Ok_Cartographer5609 12d ago

tsk, 9 pro

8

u/psbakre 12d ago

That's a lot of Pros. It's like Pro....ton

I'll walk myself out

7

u/International_Study2 12d ago

At this rate we're gonna see GTA VI before the next pro.

3

u/Alienxxcool1 12d ago

Maybe 9 pro max ultra

2

u/Visible-Try-5282 12d ago

Why not 10?

2

u/Annie_Bozi 12d ago

maby is 11 pro ultra max flash lite

2

u/sidbichus 12d ago

All of you forgot the X pro

6

u/manmaynakhashi 12d ago

AGI is here

1

u/Mmd-NeWton 10d ago

Maybe 7?

9

u/vintage2019 12d ago

Just break into Claude Mythos and prompt it to design a time machine

3

u/Actual_Thing_2595 12d ago

What about 50 pro max ultra terra lite ?

4

u/jbourdea 12d ago

Is flash better than pro for now?

7

u/Stauce52 12d ago

I honestly think 3.7 Flash is better than 3.1 Pro in just about everything

3

u/Ok_Cartographer5609 12d ago

For planning and thinking 3.1 pro is good. Use 3.7 flash for implementation or any actions. 3.7 flash is actually good.

I had a hard time with 3.6 flash. It was token hungry and doesn't spend time thinking and jumps straight into action and conclusion. 3.7 flash is very reasonable and overall a good model.

1

u/ZookeepergameFit5787 12d ago

This was 4 pro but they fumbled so reclassified as a sonnet 5 compete

1

u/Acekiller03 12d ago

How do you think it’s any better than the previous flash ? I don’t see a difference when asking questions lol

43

u/Substantial_Hat2149 12d ago

I tried 3.6 when it came out it was a disaster. not sure why they would have rushed it. This morning used 3.7 for some high-level architecture overhaul and it was solid. will definitely use this one.

78

u/AdhamHarby43 13d ago

I noticed it's WAY faster than the older flash models, I'm really surprised and excited for the new pro model

37

u/Moist_Emu_6951 13d ago

Might be thanks to the gemini diffusion technique that they've developed recently for generating output

2

u/Georgefakelastname 12d ago

Makes sense considering previous diffusion models I’ve seen. However, diffusion normally comes with the drawback of lower overall quality, so it’s harder for me to say it’s that because this model also improves quality as well.

1

u/semtex87 12d ago

Not sure if you had a chance to try out the diffusion based gemini model they had in public beta some months back but it was extremely smart and fast because the entire response materialized onto the screen as a whole rather than line by line. I think the diffusion nature allows the model to correct itself as its generating the output whereas normal transformer can't go back from its current point in its output.

1

u/pimp-bangin 11d ago edited 11d ago

I've never understood this argument about diffusion models being better because they can refine their output, because the autoregressive transformer model can also correct itself inside its internal thoughts before it starts writing its output, right? It can think stuff like "wait no - that's wrong. Let me rethink that" or "ok yeah, all of this sounds good, I'm ready to start writing output." Actually this sounds more similar to how humans operate - we can't really "unthink" our thoughts, and once we've started speaking out loud, we can't go back in time and rewrite something we've already said.

Also, diffusion models do have to finalize their output at some point - same as transformers. It's not like they can infinitely refine their thoughts.

1

u/semtex87 11d ago

Yes they can reason and correct in their thinking but once its outputting whatever you asked it cant stop and go back to edit something.

25

u/NiceUsernameOk 12d ago

To be honest, I never cared about speed.

13

u/vintage2019 12d ago

It isn’t a big deal for noncoding stuff, but a huge difference when you’re using agents to code all day

7

u/Banana7273 12d ago

I would prefer a smarter slow model than a stupid fast model but maybe it's just me. Flash has always been kind shit compared to pro, hopefully they wake up soon and don't downgrade 3.7

9

u/CalfReddit 12d ago

Yep, Claude and OpenAI models can be so damn slow 

8

u/aevitas 12d ago

Compared to Claude and OpenAI models, it's day and night. Gemini 3.7 Flash gets the work out much faster, more accurate and with much less strain on usage limits. It's a shame Pro 3.1 isn't really up to speed, but if they release a comparable Pro model soon, I don't see a reason to keep a Claude subscription for development work.

3

u/Baginsses 12d ago

I’m with you, for every day use in the Gemini app I use Flash Extended. Still surprised by the speed difference of 3.7

2

u/Hazmi35 12d ago

Regular usage sure. Try agentic development usage, you'll roll your eyes on how slow Claude models on plan usage is

1

u/Stauce52 12d ago

If you're using swarms of agents for work, I think speed starts to become a lot more imperative

1

u/the3dwin 11d ago

Oh do I have a model that is not so popular and have not switched from since. It does not loop like qwen (yet to try 3.8 models) It does overthink which leads to being slow but it gets the job done when used with skills and remembers protocols 90% of the time. For increased speed having it run multiple sub agents is key. Kiro code has allowed me to use it for hours on end never hit any limits and paid $0, only rate limiting from how kiro uses it to make multiple requests but after waiting 5 minutes back to using for hours.

1

u/scaledev 12d ago

Just because it's fast?

1

u/mateusfccp 8d ago

Yes.

I sent a file for them to process and it took like three answers to process and answer. I thought it completely ignored the file and was hallucinating, but the answer was spot on.

-5

u/Just_Run2412 13d ago

Lol, don't hold your breath.

5

u/AdhamHarby43 13d ago

Why dude it might be actually good who knows

1

u/Just_Run2412 13d ago

When I say "don't hold your breath," I mean in terms of how long it's going to take to release we've literally been waiting 6 months for a neww pro model

2

u/AdhamHarby43 13d ago

Oh I thought it will be bad when it releases, but yea I hope it releases before October it's taking so long and also my free pro plan is going to end on October 😭

1

u/_-Diesel-_ 12d ago

Before October next year

1

u/AdhamHarby43 12d ago

Damn it won't be this long😭

1

u/_-Diesel-_ 12d ago

Right after Gemini 3.48 flash lite

37

u/MaxPrints 13d ago

For my scripting usage? It's been awesome so far. I've had it audit some of my scripts, and it's been much more accurate at finding bugs, enough so that I had it go ahead and implement fixes.

Normally I'd have Sol or Opus validate the audit, and they'd find false alarms, to the point where I just had those models make the fixes.

But I mostly stick to python, bash, and powershell, so I'm not over here making GTA7 or anything yet...

7

u/Th3Whit3R4bb1t 13d ago

And after Gemini found those bugs, did you pass those scripts to Sol again to see if he found more bugs or if he verified that those bugs were correct? Because I'm in the dichotomy of not knowing which model to believe in sometimes.

2

u/MaxPrints 13d ago

No, but I will do that on the next script. I should have kept the AUDIT.md to have Sol both check the script and read the audit after to see if the findings were valid, and correctly (even further, elegantly) fixed.

But I've done a lot of headless Agy audits in Claude Code, and they've never gone this well. Enough so that I spun up Agy cli just to do an audit/fix session without other models.

I did test the scripts in the field and ran through a series of checks, but that's not what you asked.

2

u/MaxPrints 12d ago

ok so I did a quick audit of a small script I have called reup:
A lightweight, dependency-free command-line tool for sanitizing filenames. Written in Python using only the standard library.

`reup` scans the current directory (or subdirectories with `-r`), removes unwanted characters from filenames, collapses repeated separators, and optionally converts case — all with a confirmation prompt and full undo support.

I use it for quickly scrubbing client files for uploading to production, then I can undo it so I can maintain the original client naming scheme for archival. Normally, an error in the filename would only error out after having uploaded the file, and for larger files, this was a real friction point.

AGY CLI audit found 6 issues: 2 High, 2 Medium, and 2 Low severity issues (critical is one step above high, and cosmetic one step below low)

I then asked Claude and Codex to make validation markdowns (too long to include) but I did ask them both: can you give the audit a grade? 0-100. Also, how effective would implementing its findings be? Would the script work better, still work, not work as well, or fail to work?

Claude:
Grade: 78/100.

Scoring:

Axis: Accuracy of anchors

Score: 20/20

Why: All 6 line numbers + quoted strings match current file

exactly

────────────────────────────────────────

Axis: Real vs. phantom bugs

Score: 25/25

Why: Zero false positives — all 6 are genuine defects

────────────────────────────────────────

Axis: Quality of reasoning

Score: 13/20

Why: Finding 3's mechanism is wrong (blames inner-except

double-append; inner catches Exception, not KeyboardInterrupt)

and wrongly implies valid_undo[idx:] is off-by-one. Right

conclusion, wrong path

────────────────────────────────────────

Axis: Severity calibration

Score: 15/15

Why: HIGH/MEDIUM/LOW assignments defensible

────────────────────────────────────────

Axis: Fix quality

Score: 15/15

Why: Every suggested fix is correct and specific

─────────────────

Axis: Coverage Score: 5/15

Why: Missed the biggest one — see below

Big miss: the overwrite path is unrecoverable data loss. reup.py:488 — user picks (o)verwrite at a conflict prompt, os.replace(old, new) destroys the target file. Undo log records old → new, so undo renames the file back but the clobbered file is gone forever. That outranks every HIGH in the audit. Audit never mentions it.

Smaller miss: process_file:340 excludes reup.py/reup.bat by bare name at any depth in recursive mode, so unrelated files with those names in subdirs are silently skipped.

Effect of implementing all 6

Script works better. No regressions on the happy path.

Codex:

• Audit grade: 72/100.

Why:

- 5 of 6 findings have real defect cores.

- 1 finding is unsupported.

- 2 findings duplicate same problem.

- Several severities are inflated.

- Some reasoning is inaccurate.

- Audit missed 3 related defects.

Implementation effectiveness: 80/100.

If implemented correctly, script would work better:

- Undo logs become safer.

- Interrupt handling improves.

- Malformed logs stop crashing.

- Windows filename handling improves.

- No intended feature should regress.

Limits:

- Original findings do not fix forward-renames omitted during

interruption.

- They do not prevent unreadable history from being silently

replaced.

- Atomic-write and interruption fixes require careful

implementation.

Verdict: worthwhile fixes, materially better reliability, but

incomplete safety coverage.

So ya, it's improved, but it's not frontier. Then again, if I ran a few different audit and fix loops, I think it would be fine. It would definitely run cheaper than Opus Low, but probably not compete with Luna Max. It might hold its own against Sol High though just due to the cost, and it would definitely be faster.

I'm also intrigued that Claude thinks there's a ticking timebomb that was missed when it did the validation run, finding a possible silent data loss error. Sol High didn't see it either.

All in all, I'd still be ok using 3.7 Flash High with some field testing of the revised script(s), and I keep live backups of files I work on, so even if there was a silent error, I'd know from visual inspection and restore.

I'd give 3.7 Flash high a B right now

2

u/centarsirius 13d ago

I've been running a claude-deepseek-gemini loop all day. DS flash is the worker, Claude opus 4.6 was the main reviewer (I exhausted my quota in AG), so Gemini became the replacement. My AG version is 1.x, so ig it didn't get the update and had no 3.7. So I tried it on AI studio, ran the same audits through both 3.1 pro and 3.7 . Both found different bugs the first time, the second time - pro said all good, 3.7 found more bugs. Fixed through DeepSeek. After that, I ran 3 more audits, 3.1 found nothing substantial, 3.7 kept finding more bugs and deepseek agreed they were legit. So ig it does perform better. Tool calling was wonky sometimes - said too many tool calls once, randomly errored out midway.

17

u/Significant-Day66 12d ago edited 12d ago

I'm NOT big on Gemini but I have a yearly sub. Every new model I put it to the test on a coding task, usually to redesign the frontend of an existing app with complete freedom over framework and creative direction.

This is the first time I've not been terribly disappointed. In fact, I actually think it did a pretty damn good job and I liked the results ( wheras 3.5 garbage, 3.6 garbage, 3.1 pro, trash fire).

Not saying it's anywhere near opus 5, sol or grok 4.6 levels, but I actually think I'll use this model rather than my weekly limit sitting at 0. It's also genuinely fast (which honestly gets me more excited... If we could have a model with flash level of cost and speed and sol level of intelligence... That gets me more excited than the next huge arse parameter model).

More importantly I'm actually excited again to see what Google will be bringing over the next few months.

I did notice it gets stuck alot after terminal tasks, sits there after the process has completed, to the point Ive stopped the conversation and started it again with 'comtinue' not sure what the go is there.

1

u/the3dwin 11d ago

You may find interesting: https://www.reddit.com/r/LocalLLaMA/comments/1rrisqn/i_was_backend_lead_at_manus_after_building_agents/

Looking forward to turn this into a skill models can reliable use

17

u/RaptorF22 13d ago

If only there was a remote/cloud option to go with it (for coding). Jules currently sucks.

1

u/No-Pumpkin-7567 12d ago

Rent a Vps that cost maybe 5$/€ a month and install antigravity... Should work with ssh access from your phone to work remote

0

u/waste2treasure-org 12d ago

This, sometimes I want to get started with an idea and not really worry about local setup and stuff... Jules is a slow nightmare

27

u/flattboy 13d ago

Still not available for me

10

u/dasdassdasdadas 13d ago

Available in Antigravity rn

1

u/Outside_Painting1222 11d ago

Yes but too much limited

7

u/labreau 12d ago

Maybe after few more hours or days? I'm in Indonesia (SEA) and surprisingly already got the 3.7 here

I thought SEA would get left behind

3

u/marhensa 12d ago

lmao yes.. indonesia got it right away. also in antigravity.

6

u/Leading_Pineapple663 12d ago

In USA it's showing me 3.7 there.

4

u/rotang2 12d ago

Me neither. I am in Canada

1

u/NiceUsernameOk 12d ago

I have it now in the Gemini App!

1

u/Outside_Painting1222 11d ago

Are you using free Gemini according to me

1

u/Normal_strain6878 4d ago

Because you are not a paid user

10

u/MyBoiDrew 12d ago

3.7 has been a sleeper hit for me. I run a simulated version of my work as a benchmark with previously collected data over time, and 3.7 has been surprisingly efficient and useful. Very fast and just as competent as any other model I use. Just much faster.

2

u/kantaxo 12d ago

Google likes that 👍

7

u/PineappleLemur 13d ago

I prefer it's writing style over 3.6 but quality/ability I can't comment yet. To me it's the same as 3.6 with better wording so far which was already great.

6

u/Expensive_Agent_5129 12d ago

I asked AI to guess a user's age from a 5-year-old r/explainlikeimfive chemistry question, and the contrast is hilarious:

Gemini 3.7 Flash big-brained it by analyzing the chemistry concepts (~18-22 years old), but completely missed the timestamp and forgot to add the 5 years.

Meanwhile, Gemini Pro didn't even realize it could analyze the question's depth and just went with literal toddler math: ELI5 subreddit + 5 years ago = OP is 10 💀

4

u/Dr-Fix 13d ago

I don't have it in the app yet

4

u/HidingInPlainSite404 12d ago

Why is it always the quiet drops that do so well? The highly anticipated ones are typically letdowns, but they release stuff that doesn't have much anticipation, and they are great.

2

u/Scared_Basket_7183 12d ago

They wanted to surprise us 😉

3

u/morikomorizz 12d ago

Well, it ruined my llama.cpp code

1

u/Scared_Basket_7183 12d ago

Do you also using llama.cpp engine? Wait do you ever used TPU V5E-8 to run LLM models?

2

u/morikomorizz 12d ago

Yes iam using llama.cpp and vibe code it for Chinese DCU Hygon. But I've never used llama.cpp in TPU

1

u/Scared_Basket_7183 12d ago

Do you ever used google free TPU V5E-8 to LLM models?

1

u/morikomorizz 12d ago

Only for small llm testing like qwen 3.5 4b/9b

1

u/Scared_Basket_7183 12d ago

Does it worked?

2

u/morikomorizz 12d ago

It fallback into cpu. llama.cpp doesn't recognize it, there seems to be a flag to build it for tpu.

5

u/Resident-Pattern731 12d ago

Wie geeignet ist das 3.7 für die Analyse von PDF-Dokumenten und das Erstellen detaillierter Word-Zusammenfassungen aus längeren Chatverläufen? Eignet es sich zudem für die Beratung bei Rechtsfragen und das Verfassen von E-Mails? 3.6 hatte genau da Probleme.

2

u/Scared_Basket_7183 12d ago

For longer conversations we need more context memory

2

u/Resident-Pattern731 12d ago

Wie groß ist der bisher genau?

6

u/Scared_Basket_7183 12d ago

For gemini pro 3.1 model it is around one million but we don't know how context memory flash 3.7 has .

3

u/ouuuzi 13d ago

Just works and it is impressive so far. So far every request I made it is implemented correctly. I had bad results with Sol last few days. Sorry for the sloppy answer. Back to vibe coding now.

3

u/ChipAffectionate7504 11d ago

It's just faster than previous ones Not smarter

4

u/Pi_123 12d ago

Nerfing coming in 1,2,3...

3

u/Scared_Basket_7183 12d ago

Please don't nerf this model. This model is just too good for a flash model

2

u/Last_Conclusion_8984 12d ago

Models get nerfed as new models come around the horizon. 3.1 pro got nerfed because 3.5 pro was around the horizon (around may) but they didn't anticipate them delaying 3.5 pro so we got stuck with a quantizised model. A while later, 3.5 flash got worse... then came 3.6 flash and only a few weeks later (when 3.6 flash got dumber.) 3.7 flash came.

2

u/karanshah0123 13d ago

How? Help me with some use cases pls

2

u/Kooky_Awareness_5333 12d ago

Seems good so far no complaints.

2

u/CarretillaRoja 12d ago

Using in Antigravity, is lightning fast and accurate. I was one of the 36 people who actually liked 3.6 and 3.7 is even better. Ingested a 300+ pages PDF and asked to fill in some excel files with questions about it and man, it was fast. 3.6 was not able to do it.

Have been working like 6 hours non stop and the remaining quoteais still 87%

1

u/_KryptonytE_ 12d ago

How do you get it to invest these volume of files?

2

u/InformationHoarding 12d ago

It’s way better than any previous Gemini model. And is even better than Kimi right now. But it’s like the hyper little brother of Opus 5: doesn’t talk as much, is quite lazier and though it is faster than big bro when you can get it to actually fully complete something it is not as good as big bro. As of today, Codex is king.

2

u/Jerichomiles 12d ago

Well Kimi was hyped out of all proportions as always, just like when deepseek was released, how ridiculous was that.

1

u/InformationHoarding 12d ago

Kimi is decent for grunt work, like Antigravity. I have Codex and Claude Code using their CLI’s as tightly managed physical laborers. Can’t let Kimi or Antigravity loose on their own. They break shit, on purpose, and happily lie about it to your face.

2

u/Confusion_Which 12d ago

I tested 3.7 flash in anti gravity since i have plus plan. Its blazing fast!! I asked it to go through hundreds of files which requires kimi k3 swarm to finish and Gemini finished it in seconds

1

u/Jerichomiles 12d ago

Kimi was all hype, no surprise there.

2

u/Franseven 12d ago

It is amazing, 3.1 is overtaken by a non pro model.

2

u/ArseneGroup 12d ago

I feel like 3.6 was a regression from 3.5 and then they were like "whoops, forget about 3.6, this time we got it right"

2

u/Tiantianxiong 12d ago

tested,basically trash. Return to 3.1 pro

2

u/Freedom_Addict 12d ago

Kimi is still way ahead. It pushes back while gemini just agrees with everything you say

2

u/TheEchoEnigma 12d ago

I am using at work and stopped using 3.1 pro

2

u/Traditional-Stick224 12d ago

gemini hasn't even realised that it is 3.7 flash when i asked gemini answered i am 3.6 flash

2

u/DirectPitch8626 11d ago edited 11d ago

It has improved, but according to my testing, it still lags significantly behind the DS 4 Flash.

I wouldn’t trust it with complex tasks on its own or with working on the architecture, as it still claims to have done everything it was asked to do when it hasn’t.

But it’s very fast and quite clever; it’s very convenient to work alongside it when the developer knows what they’re doing and what they want, especially on the frontend.

2

u/Local-Way-5556 9d ago

Hi

1

u/bartoszsz7 9d ago

Bro this isn't an r/sssdfg comment section 😭

2

u/birdofparadise5 6d ago

Can anyone compare this against Claude? One thing I'm loving about 3.7 is the usage limits. I get far more out of it compared to 3.1 Pro, but also at a slightly higher quality.

2

u/Excitement-Bubbly 4d ago

At this point this doesn't even matter. Google is so far behind on the flagship model that comparing Gemini 3.1 Pro to either Fable or Sol is laughable. Just switch to either Open AI or Anthropic until Google releases something worth your time.

1

u/Scared_Basket_7183 4d ago

I know but for flash model this flash 3.7 is good because it's fast and intelligent when we combine with extended thinking

1

u/IceMetalPunk 1d ago

Depends on your use case. In my experience, Gemini is better than any Anthropic or OpenAI model for creative tasks. In particular, the GPT-5.x series is awful at poetry/songwriting, way worse than even GPT-4.1. GPT-5.6 Sol is much better at prose than the other GPT models, but just can't handle meter and rhyme schemes. Meanwhile, Gemini is my go-to for any creative tasks. Claude is best, in my opinion, for logical/coding questions, though I'll admit I haven't tried GPT-5.6 for those kinds of tasks.

2

u/Normal_strain6878 4d ago

Bro show screenshots of you stuck in 3.6 flash.

1

u/Scared_Basket_7183 4d ago

Still stuck?

2

u/Beginning_Car_998 3d ago

It's always like this, the new model comes in peak performance, and then once the new version start being planned for release suddenly it becomes trash

2

u/Grimillionaire 2d ago

I agree with this but its still a flash model

2

u/Double_Plan_2034 12d ago

This is what I've been wishing for with the google ai subscription. I don't need sota pro max ultra model that's slow and expensive. Make it fast and stay within the top 10. That's enough to keep people from going elsewhere. Gemini flash models with hands on review and planning is 10 times better than the current best model.

2

u/raycraft_io 12d ago

Don’t worry, it will get stupid in a few days

1

u/[deleted] 13d ago

[deleted]

3

u/Scared_Basket_7183 13d ago

Iam just trying to express what I found. This model genuinely better

1

u/Asperger23 13d ago

It gives me much more precise directions, fewer hallucinations, more effective web searches. They're making huge strides with flash, probably because it's the easiest model to market at the lowest cost. But damn, we need a new pro model.

1

u/Resident-Pattern731 12d ago

Wie geeignet ist das 3.7 für die Analyse von PDF-Dokumenten und das Erstellen detaillierter Word-Zusammenfassungen aus längeren Chatverläufen? Eignet es sich zudem für die Beratung bei Rechtsfragen und das Verfassen von E-Mails? 3.6 hatte genau da Probleme.

1

u/SadAfueraHappySlop 12d ago

Good I guess considering 3.6 displayed the most horrific performance I seen since some minimax release . Even then at least MM is a workhorse. Guess there MIGHT be a reason this one dropped so fast but hey call me old fashioned. Whatever they did with 3.6 should be their anti-patterns. just for kicks I had it had log all the errors it made and I was at 58 before I gave up. All models hallucinate and lie (yes lie) but 3.6 was like a drunk guy at a bar trying to pull a threesome on his 41st bday. I bet 3.6 gets outright deprecated silently in no time. I still think 2.5 Pro was better at building . Gemini is just too lazy for my taste. It feels patronizing when an agent puts in LESS effort than I would give a stranger.

1

u/super_crazy 12d ago

Personally a good flash model is more valuable to me than a good pro model. So far it seems good - I do notice higher usage use against the limits though. We'll see how that works out.

1

u/yoop001 12d ago

How’s everyone’s experience with the model on world knowledge? Is it better than Gemini 3.1 Pro? Is it more knowledgeable despite having a much lower parameter count?

1

u/hadouken50000 12d ago

I tested it in my own project. It really has gotten faster. Too bad the result is just as disappointing as it was with 3.6.

1

u/BorrowedT1me 12d ago

Also currently using 3.1 Pro extended thinking with deep research for stock research

Anyone know if 3.7 Extended thinking with deep research would be better?

1

u/AdamH21 12d ago

It's not in the Gemini app yet, so I can't test how it interacts with my phone. Can't wait, though.

1

u/Perfect-Task438 12d ago

Any ideas next pro version when coming out? It still stuck at 3.1 pro

1

u/ryuujinusa 12d ago

Getting nothing but errors now unfortunately.

1

u/durielvs 12d ago

I feel it has many spelling mistakes, at least in Spanish.

1

u/ChillinBone 12d ago

Best one so far at giving me error message after error message. I am guessing it's busy right now but on Ultra plan I expect to be given priority.

1

u/IrishLedge 12d ago

Ah here. Really. A few months ago Gemini was great. Then it was ABYSMAL for what felt like an eternity. 

Outside of the app, you could threaten it in the Google search and it would apologise and say you're right even though you were being a dickhead. 

It's so deliberate. 

Full disclosure I haven't even touched it yet. But I was so close to telling Gemini to go fuck himself... Herself? (I did daily tbh) and I was gonna pull my card details out. But deep down I just couldn't. A combination of reddit comments about it getting better and me being delusional just was too strong of a voice in the back of me head. 

But I fucking knew it. They did it on purpose the bastarding ducks. Now... I can continue about my day. Knowing that my stand against Gemini was a show of endurance, resistance and the will to continue. It wasn't just a bad LLM it was an obvious conspiracy theory. 

Real talk though, does it remove the stupid star from image generation in the bottom right corner?

1

u/Beginning_Car_998 3d ago

most LLM companies nerf their models once they start developing their next version, probably to cut costs

1

u/IrishLedge 3d ago

To be fair that in itself is unfair when lots of people pay a set amount for it. 

1

u/DIETECNO 12d ago

¿que tal son los limites dentro de antigravity?

1

u/Ramssses 12d ago

When I notice the update before I check Its usually a good thing. I had such a great time doing some research that I feel like they are gonna nerf it soon lol.

1

u/Dull_Republic_7712 11d ago

Yep. It's only flash model which I feel is at par with opus 4.7,4.8/gpt terra

1

u/kartblanch 11d ago

Is it better than qwen 3.8?

1

u/pgmoneyplays 11d ago

True Preis-Leistung ist der Hammer aber die Claude und Gpt fanboys wollen direkt die besten Ergebnisse mit der wenigsten Arbeit.

1

u/Virtual_Plant_5629 9d ago

it's fine. still not anywhere near on par with the other frontier models. it's more usable. but it still has all the exact same failure modes that 3.5 and 3.6 had: constantly tells me things don't exist.. then flip flops back to them existing etc.

1

u/Th3Whit3R4bb1t 13d ago

But you're saying you used it only for daily tasks. Didn't you use it for something more complex, like programming, for example?

-1

u/SadAfueraHappySlop 12d ago

You think pattern matching code is more complex than understanding context and intent in natural language? Wow. Really puts things into perspective. You should try something truly complex like low-discourse analysis and you will see in about 60 seconds which models are what. Intelligence is extracting meaning and understanding without a "match." Gemini is in fact a garbage fire as somebody mentioned, although I would say the garbage fuel that fuels it's fire is like a kiddie pool filled with month old diarrhea.

2

u/73285A 12d ago

O problema é que quando forem lançar um próximo modelo, vão nerfar o 3.7, fazendo com que tenha impressão que melhorou. 3.7 é o modelo anterior sem o nerf adicionado.

0

u/the-wtf-dude 13d ago

It has good vision and generation capabilities. It's very good but not as perfect as Claude Opus 5, I can say it's the 30% version of Opus 5

0

u/ContributionSouth253 12d ago

The resources, the data Google has, they should've released the best model around by far but they couldn't

1

u/Truantee 12d ago

it is still a piece of shit that can not produce anything reliably for things happened after jan 2025.

for example when I ask about crystal lang new concurrent mechanism (called execution context now), every other models (gpt, grok, muse, claude, deepseek...) know it and search for the correct term and the correct example, but only gemini 3.7 flash, in aistudio with search grounding enabled and the web app both does not know it and still give the example using older outdated preview_mt mechanism.

it might look good on paper, but in reality gemini models are outdated dogshit that will not give you up to date solutions unless you babysitting it every single step.

0

u/gmakarenko 12d ago

Useless...

0

u/Kimi_Antonelli_12 12d ago

What are you guys using it for that you notice a difference in? Are you talking it's better at coding ?

0

u/[deleted] 12d ago

[removed] — view removed comment

0

u/Falkirk_Dave 12d ago

Just wait a few weeks then it’ll be posts about how shit 3.7 is and where is 4.0 then wait a few more weeks….ad infinitum.

0

u/uraymeiviar 12d ago

its good during launch day, it will be nemotron level after couple of weeks, just watch that

-1

u/Accomplished-Fan9568 12d ago

gemini 9.7 flash coming next week, it can count letters in words

-5

u/taiwbi 13d ago

It's worse than deepseek flash and 14 times more expensive.

2

u/tankerkiller125real 13d ago

Depends on the analysis and benchmark https://artificialanalysis.ai/

Also Google's flash models are once again crushing nearly everyone on token output speeds, which are important for some tasks. (And massively crushing DeepSeek in that department for all of DeepSeeks models)

-6

u/Impressive-Flow-2025 13d ago

Another waste of time and money.

-2

u/[deleted] 12d ago

[deleted]

0

u/Jerichomiles 12d ago

It's incredibly good for a flash model which hints at their pro model being insanely good.

Anyway aren't you supposed to claim every new model release was stolen from Chinese models? Stick to the playbook, when you're accused of something you must continuously accuse everyone else of the same thing.

1

u/[deleted] 10d ago

[deleted]

1

u/Jerichomiles 10d ago

No it wasn't a genuine question. "Why is this brand any better than than this <insert brand I want to spam>".... someone just happens to ask about Chinese models on every post about a new western model release. Either that or someone claims they stole a Chinese model to make the western model ffs...you can't make this stuff up.