"Do X", "X is done!", "Are you sure?", "Yes!", "Did you test it?", "Yes!", "It doesn't look done and nothing works.", "I may have overstated completion. It is 7% done."
This is getting old.... Anyone else seeing this with Sol 6 AND 6.1? Even if I create a plan with numbered objectives, a literal checklist, it still regularly overstates the amount of work has been done. Any tricks to keeping it on task and not outright lying about its progress? EDIT: "I followed th…
Read the full story at r/OpenAI ↗
Timeline · 1 report
- 2026-10-02 01:07 · r/OpenAI
"Do X", "X is done!", "Are you sure?", "Yes!", "Did you test it?", "Yes!", "It doesn't look done and nothing works.", "I may have overstated completion. It is 7% done."
More stories
- Bring near-Astra intelligence to everyday work with GPT-6.1 Sol on Amazon Bedrock — AWS Machine Learning Blog
- Gemini 4 Argon: our next era of frontier intelligence — Google Gemini Blog
- Introducing Olmo-core 3: Open, scalable training infrastructure for large MoEs — Allen Institute for AI (Ai2)
- Google Releases New Gemini Model With Guardrails Amid A.I. Safety Debate — New York Times Technology
- OpenAI postpones release of latest AI model over security concerns as the industry faces new safety pressures — Euronews Next
- Introducing GPT-6.1 Sol — OpenAI News
- Google tests its plan for AI data centers in space with Project Suncatcher — Scientific American
- OpenAI’s Dots Are Always-On AI Agents—and Its Answer to Meta’s Muse — Wired AI
Get the daily brief of stories like this at 6:30 every morning →