For almost every section in the model card there is the message: Gemini 3.7 Flash is based on Gemini 3.6 Flash.
Same training dataset, same software, same hardware, same architecture...
I'm wondering what they changed actually for the model to be more powerful if the benchmark results are real and relevant.
Maybe just tweak settings or the reasoning prompts and called it a new version of their model?