AINewsnow

Claude Opus 5.5 placed 3rd of 14 models at rebuilding photos in Blender, 1 point off first on the desk. Plus a caching mistake worth knowing about

I'm building a photo-to-Blender tool and benchmarked 14 models as its agent: look at a photo, write and run Blender Python, render, compare, repeat. Hard caps of 20 minutes, $4 and 60 requests per scene. Scoring is deterministic code, not an LLM judge. How Claude did (average of three photos, 0–100…

Read the full story at r/ClaudeAI ↗

Timeline · 1 report

  1. 2026-10-02 10:31 · r/ClaudeAI
    Claude Opus 5.5 placed 3rd of 14 models at rebuilding photos in Blender, 1 point off first on the desk. Plus a caching mistake worth knowing about

More stories

  1. Build a multi-agent music production pipeline on Amazon Bedrock AgentCore Runtime Instances — AWS Machine Learning Blog
  2. Google unveils Gemini 4 Argon with SOTA score on DeepSWE — TestingCatalog AI News
  3. Introducing Anthropic models on Amazon Bedrock for in-region inference in Seoul and Singapore — AWS Machine Learning Blog
  4. Kimi K3: A Claude clone or something else? — CoreWeave Blog
  5. My tldr for OpenAI dev day — r/singularity
  6. OpenAI's new GPT-6.1 Sol undercuts its own Astra flagship — The New Stack AI
  7. Implementing Multi-Environment Access for Claude Platform on AWS — AWS Machine Learning Blog
  8. Amazon Bedrock expands Claude model availability to in-country inferencing in India — AWS Machine Learning Blog

Get the daily brief of stories like this at 6:30 every morning →