Gemini 3.7 Flash: 87/98 vs 74/98 for 3.6 Flash, while also getting ~40% faster
This story is from 2026-08-26. It is preserved in the archive; the latest stories are on the live feed.
I ran Gemini 3.7 Flash on the current 98-task MindTrial set with high thinking and the same Python executor used for the other models. The improvement over Gemini 3.6 Flash was much larger than I expected: Gemini 3.6 Flash: 74/98, 1 hard error, ~1h45m Gemini 3.7 Flash: 87/98, 0 hard errors, ~1h03m…
Read the full story at r/GeminiAI ↗
Timeline · 1 report
- 2026-08-26 22:47 · r/GeminiAI
Gemini 3.7 Flash: 87/98 vs 74/98 for 3.6 Flash, while also getting ~40% faster