logoalt Hacker News

Gemini 3.7 Flash

176 pointsby thisisauseridtoday at 5:23 PM123 commentsview on HN

https://ai.google.dev/gemini-api/docs/models/gemini-3.7-flas...


Comments

jjcmtoday at 6:06 PM

Here's a image->html test. Gemini has always swung above its weight class for vision work, so I'm always eager to try it with this.

Original images: https://image.non.io/neonRamenDesigns.webp

Gemini 3.7 build: https://html.non.io/neonRamenGemini3.7

Opus 5 build for comparison: https://html.non.io/neonRamen

Opus is still best in class for this, but it's worth noting how well Gemini 3.7 does vs a more comparable LLM price wise, which is Grok 4.6: https://html.non.io/neonRamenGrok4.6 . I thought Gemini would blow Grok out of the water (it generally has in the past), but Grok has really caught up.

show 1 reply
eckrtoday at 6:16 PM

Maybe this is just my experience, but have people had trouble with 3.6 Flash just... getting things it has seen in its context correct? I don't know if it's been insanely benchmaxxed or what, but it'll pull information from websites and immediately get it wrong the token after. Or for example (this is something that happened like yesterday) I asked it to compare the uses of A and B in a language I was learning, and the way I typed it was "Please compare how these two are compared differently: A VS B", and then... it proceeded to compare "VS" and "B". I'm not kidding.

Personally whenever I use Gemini I've just been using 3.1 Pro because I've had insane trouble with them getting things incorrect like this. Hopefully they'll fix it soon / they've fixed it with 3.7 Flash.

wxwtoday at 5:37 PM

They need to release benchmarks against Luna/Terra. Luna is much cheaper which feels like it undercuts the need for Flash.

I've always considered the Flash series of models to be for low-cost, high-volume, mostly text-based use cases (e.g. summarization, parsing, formatting), emphasis on low-cost.

[edit: ah, benchmarks here: https://blog.google/innovation-and-ai/models-and-research/ge...

more of a Terra than Luna competitor which is an interesting positioning. I feel like differentiation at the mid-tier of models is pretty difficult.]

show 5 replies
euazOntoday at 5:28 PM

The multimodal abilities are great, but if you deal with text only, what is the benefit of using this over DS V4 Flash/Pro? 13-26x cheaper with comparable intelligence, and available across many different inference providers.

I fail to see the usecase where DS V4 Pro is not enough, but Flash 3.7 is - except multimodal.

Luna is similar, and also 8x cheaper. Source: artificialanalysis

The only benefit I can see is the speed, that looks to be outstanding, probably thanks to their TPUs.

show 5 replies
twelvechairstoday at 5:36 PM

https://artificialanalysis.ai/models/gemini-3-7-flash

The selling point for gemini continues to be speed and particularly end-to-end response time.

show 4 replies
ghoshbishakhtoday at 6:15 PM

Has anyone noticed that antigravity has been working really well for the last few weeks. Now with this model it should be working much better. Hope the Google AI Pro Subscription can be used to do some real agentic coding now.

parastitoday at 5:50 PM

Actual announcement: https://blog.google/innovation-and-ai/models-and-research/ge...

So it's better than 3.6 Flash, at half the price. I've been pretty excited about Gemini models recently, they just feel so fast after spending most of the day at work waiting for Opus 5.

show 2 replies
fmind-devtoday at 5:34 PM

Gemini Flash is one of the best "good-enough" models. I use this type of model daily, for automation and quick development iteration loops.

Unfortunately, it's often not strong enough for heavy refactoring and long running development loops.

show 1 reply
Topfitoday at 5:31 PM

> What's new in Gemini 3.7 Flash [0]

> Coding and agentic tasks: Significantly higher quality on real-world software engineering and agentic benchmarks, improving issue resolution and reducing failed agent loops.

> Web development and stronger design parity: Generates higher-fidelity desktop and web application code directly from design mocks, with strong gains in design adherence and in auditing existing codebases against mocks to verify 1:1 design parity.

> Promotional pricing: Gemini 3.7 Flash will be available at an introductory price of $0.75/1M input tokens and $3.75/1M output tokens. We’re also applying this new rate to 3.6 Flash. Introductory pricing expires on December 31, 2026; after, $1.50/1M input tokens and $7.50/1M output tokens will apply.

Still no sign of 3.5 Pro. Will have to test it, low expectations given every other model from the Gemini 3 lineage, but one can hope. Just struggle to understand the promotional pricing being temporary for four months. Given this industry, I'd be hard pressed if 3.7 Flash was still in use by end of year, so why not make it the official pricing?

[0] https://ai.google.dev/gemini-api/docs/latest-model

show 1 reply
impulser_today at 6:11 PM

After being stuck with using GPT-5.6 models for the past few weeks, I have renewed faith in Google and everyone but OpenAI. The GPT-5.6 models are quite obviously benchmarkmaxxed to make they seem like they are intelligent but they are quite dumb outside anything that not a benchmarked task.

I also think Google is still the best at fitting the most overall intelligences into their models, but for some reason it seems like the model architecture is just bad.

Alifatisktoday at 6:09 PM

Ever since the insane discount with GPT-5.6 Luna, not much excites me anymore. I mean just look at the benchmarks, even though Gemini 3.7 Flash performs well on the DeepSWE 1.1, Luna (Max) still performs way better.

https://deepswe.datacurve.ai

> Starting January 1, 2027, $1.50/1M input tokens and $7.50/1M output tokens will apply.

Compare this to Luna which is at $0.2/1M input ($0.02 cached) and $1.2/1M output.

https://developers.openai.com/api/docs/models/gpt-5.6-luna

damstatoday at 5:30 PM

> 3.7 Flash is available through the end of the year at an introductory price of $0.75/1M input tokens and $3.75/1M output tokens.

> Introductory pricing expires on December 31, 2026. Starting January 1, 2027, $1.50/1M input tokens and $7.50/1M output tokens will apply.

show 1 reply
bisonbeartoday at 5:26 PM

Reposting my comment from the other thread https://news.ycombinator.com/item?id=49288847

They compare it to 5.6 Terra, however https://cognition.com/frontiercode puts Terra at about 1/2 the price

Also have to compare to the recent Grok 4.6 release, which appears to straight up be better AND cheaper

Hard to understand why anyone would choose 3.7 Flash under these conditions.. is Deepmind still a frontier lab?

show 1 reply
bisonbeartoday at 5:25 PM

They compare it to 5.6 Terra, however https://cognition.com/frontiercode puts Terra at about 1/2 the price

Also have to compare to the recent Grok 4.6 release, which appears to straight up be better AND cheaper

Hard to understand why anyone would choose 3.7 Flash under these conditions.. is Deepmind still a frontier lab?

show 3 replies
Tiberiumtoday at 5:28 PM

3.7 Flash gets 56 on AA up from 52 for 3.6 Flash. But it seems like this is at the cost of more output tokens per task: 3.6 Flash is 26k, 3.7 Flash is 37k. Due to 3.7 Flash's 2x slashed pricing it's still cheaper per task.

nickandbrotoday at 5:27 PM

This is genuinely a competitive model, considering it beats Claude Sonnet 5 on almost all benchmarks and is more than half its price. Seems like Google is back in the game, though not leading the frontier anymore.

show 3 replies
cracadumitoday at 5:37 PM

For those looking for the full benchmark figures and technical overview, Google's primary announcement post is here: https://blog.google/innovation-and-ai/models-and-research/ge...

npntoday at 5:30 PM

> * For 3.6 and 3.7 Flash, introductory price expires on December 31, 2026. Starting January 1, 2027, $1.50/1M input tokens and $7.50/1M output tokens will apply.

this is hilarious. it is not 2025 any more, by Jan 2027 there will be at least 3 newer generation of models (from other provider) released already. nobody would use flash 3.7 at that time.

sure we used to cling to gemini models in the past, demanding 2.5 models to continue to serve, but since google betrayed us with those price hike, people already spent their time making their production pipeline less dependent on google since then.

heck, even now I'm not sure I even care if they cut the pricing even lower. there are too many models with cheaper price and similar performance now.

show 7 replies
spelktoday at 5:07 PM

>3.7 Flash is available through the end of the year at an introductory price 1 of $0.75/1M input tokens and $3.75/1M output tokens. This price combined with the enhanced model performance enables developers and customers to scale production-ready agents cost effectively.

Introductory pricing until December 2026 implies no significant Gemini Flash developments until the next year.

show 2 replies
stillpointlabtoday at 5:39 PM

Does Google believe people want fast models because they have some sort of evidence of that preference? Or are they no longer capable of delivering a Pro model?

show 5 replies
nateb2022today at 5:36 PM

[dupe] https://news.ycombinator.com/item?id=49288847 (35 points, 8 comments)

rodolphoarrudatoday at 5:46 PM

Did the company fix the high friction between any service and their models' API?

I hope so. It seems mind boggling to me that an user needs to surf around different sections (plural) of google cloud console, then this Vertex and do a dozen clicks to issue a simple key.

show 1 reply
orliesaurustoday at 5:30 PM

what a week - lets see it draw a weird animal doing a weird thing on a bicycle

show 1 reply
nomilktoday at 5:43 PM

How does it compare to Opus 5.0 and Fable 5 for coding? E.g. in Cursor or OpenCode?

show 1 reply
bartmantoday at 5:39 PM

At the discounted rates, upgrading from 3 Flash to 3.7 Flash is finally reasonable.

In my evals 3.6 Flash (pre price change) was usually a bit more token efficient than 3 Flash, so I‘m expecting same or even lower cost-per-task on 3.7.

Maybe a play by Google to deprecate 3 Flash soon.

algoth1today at 5:37 PM

Well, you do get 1 million tokens and the ability to reason over video natively and many of us are forced to pay for 20usd plan anyway due to google drive 5TB, not to mention notebooklm, so it’s not a nothing burguer, it’s just an almost nothing burguer

pkoirdtoday at 5:27 PM

When are we getting another pro model from Gemini? Or are they simply focusing on the niche of fast but moderately capable models?

khanhnguyen8386today at 5:37 PM

Offering a 'temporary introductory discount' until Dec 2026 on an LLM is hilarious. In this market, by Jan 2027 this model will be superseded by 5 different providers offering 10x the performance at half the post-discount price anyway.

9cb14c1ec0today at 5:27 PM

Model card: https://deepmind.google/models/model-cards/gemini-3-7-flash/

Somewhere in the same neighborhood as GPT 5.6 Tera and Sonnet 5, depending on the bench.

jespineltoday at 5:54 PM

IMO, they should drop their previous model (3.6 Flash) from the benchmark charts. I don't care how better this is compared with their previous model. What matters (to me) is:

1. How the new model performs against the other top models in the same category.

2. The pricing of the new model against the other top models in the same category.

cmrdporcupinetoday at 6:06 PM

So, again with a Flash model. Why are they so afraid to put out an actual SOTA frontier high intelligence model?

We still don't have a 3.5 Pro, and along comes 3.7 Flash?!

greatgibtoday at 6:06 PM

For almost every section in the model card there is the message: Gemini 3.7 Flash is based on Gemini 3.6 Flash.

Same training dataset, same software, same hardware, same architecture...

I'm wondering what they changed actually for the model to be more powerful if the benchmark results are real and relevant.

Maybe just tweak settings or the reasoning prompts and called it a new version of their model?

yanis_ttoday at 5:31 PM

Is that he model that supposed to be Pro, but then they changed their mind?

show 1 reply
keketitoday at 5:46 PM

In August of 2026, Gemini became self-aware, and began producing increasingly crappy flash versions of itself...

brendongtoday at 5:45 PM

Glad to see that the company with the most data is releasing the most amount of models. Some things do make sense

ghoshbishakhtoday at 6:02 PM

So a bit worse than DeepSeek.

TekMoltoday at 5:32 PM

I'm only interested in the state-of-the-art model by each provider.

For Google, this is still gemini-3.1-pro-preview, right?

show 2 replies
toshtoday at 5:35 PM

strong improvement over 3.6 flash

but luna is hard to beat @ capability / cost

eistoday at 5:46 PM

Grok, Meta, Gemini and others all released updates to their models within around a month or two from their respective last release and made significant jumps in benchmarks all around the same time. Any guesses as to why that is? Is it just the release season and/or everyone is benchmaxxing?

show 1 reply
jdw64today at 5:36 PM

I'm really curious about this: the foundational paper behind today's LLMs came from Google, and some of the world's best scientists were at Google. So why are they falling so far behind in the AI race?

show 1 reply
AntonioEritastoday at 5:27 PM

Another failed 3.5 pro run branded as 3.7 flash. It's getting sad.

show 1 reply