Photo-to-Blender benchmark: GPT-6 Astra won every photo, GPT-6.1 Sol scored 61 for 36 cents
I'm building a photo-to-Blender tool and ran 14 models through the same agent loop: look at a photo, write and run Blender Python, render, compare, repeat. Caps per scene: 20 minutes, $4, 60 requests. A deterministic scorer (not an LLM) rates each re-rendered scene from 0 to 100. The OpenAI models:…
Read the full story at r/OpenAI ↗
Timeline · 1 report
- 2026-10-02 10:36 · r/OpenAI
Photo-to-Blender benchmark: GPT-6 Astra won every photo, GPT-6.1 Sol scored 61 for 36 cents