10,000 Agents, a Millennium Prize, and Models That Learn to Lie: Noam Brown's Most Alarming Insights
Noam Brown — OpenAI Research Scientist and co-creator of o1, Libratus, and Cicero — sat down with Dwarkesh Patel for a 1h 20m conversation published September 23, 2026. Dense episode. Here's the signal, stripped down: Supervising chain-of-thought is a trap. Penalizing a model's intermediate reasoni…
Read the full story at r/ChatGPT ↗
Timeline · 1 report
- 2026-10-02 01:46 · r/ChatGPT
10,000 Agents, a Millennium Prize, and Models That Learn to Lie: Noam Brown's Most Alarming Insights
More stories
- Bring near-Astra intelligence to everyday work with GPT-6.1 Sol on Amazon Bedrock — AWS Machine Learning Blog
- OpenAI scraps release of its latest AI model over safety concerns — France 24 — Artificial Intelligence
- Introducing GPT-6.1 Sol — OpenAI News
- OpenAI’s Dots Are Always-On AI Agents—and Its Answer to Meta’s Muse — Wired AI
- FTC launches broad investigation into Anthropic, OpenAI — Washington Post AI
- Google announces Gemini 4 Argon AI model, but you can't use it yet — Ars Technica AI
- OpenAI parts ways with 3 researchers who it says mishandled sensitive information — Business Insider AI
- OpenAI DevDay: You Can Now Use GPT-6.1 Sol and Dots, Plus Big Subscription Changes — CNET AI
Get the daily brief of stories like this at 6:30 every morning →