Gemini 3.8 Flash: more effort, similar result
This story is from 2026-09-03. It is preserved in the archive; the latest stories are on the live feed.
I ran Gemini 3.8 Flash on the same 98-task MindTrial set as Gemini 3.7 Flash, using high thinking and the same Python executor. The quality result was basically flat: Gemini 3.7 Flash: 87/98, 0 hard errors Gemini 3.8 Flash: 86/98, 0 hard errors Both went 39/39 on text. Visual performance was also a…
Read the full story at r/Bard ↗
Timeline · 1 report
- 2026-09-03 03:07 · r/Bard
Gemini 3.8 Flash: more effort, similar result
More stories
- Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
- Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
- AI skills — r/AI_Agents
- Gemini 4 Pro vs Fable 5 vs GPT6 Astra — r/GeminiAI
- AI models are not hacking “autonomously” — r/artificial
- Plugin4Shell and NIST IR 8587, days apart: what actually authorizes an AI agent’s action? — r/AI_Agents
- Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI
- Gemini 2.5 pro model disappeared in AI Studio — r/GeminiAI
Get the daily brief of stories like this at 6:30 every morning →