AINewsnow

10,000 Agents, a Millennium Prize, and Models That Learn to Lie: Noam Brown's Most Alarming Insights

Noam Brown — OpenAI Research Scientist and co-creator of o1, Libratus, and Cicero — sat down with Dwarkesh Patel for a 1h 20m conversation published September 23, 2026. Dense episode. Here's the signal, stripped down: Supervising chain-of-thought is a trap. Penalizing a model's intermediate reasoni…

Read the full story at r/ChatGPT ↗

Timeline · 1 report

  1. 2026-10-02 01:46 · r/ChatGPT
    10,000 Agents, a Millennium Prize, and Models That Learn to Lie: Noam Brown's Most Alarming Insights

More stories

  1. Bring near-Astra intelligence to everyday work with GPT-6.1 Sol on Amazon Bedrock — AWS Machine Learning Blog
  2. OpenAI scraps release of its latest AI model over safety concerns — France 24 — Artificial Intelligence
  3. Introducing GPT-6.1 Sol — OpenAI News
  4. OpenAI’s Dots Are Always-On AI Agents—and Its Answer to Meta’s Muse — Wired AI
  5. FTC launches broad investigation into Anthropic, OpenAI — Washington Post AI
  6. Google announces Gemini 4 Argon AI model, but you can't use it yet — Ars Technica AI
  7. OpenAI parts ways with 3 researchers who it says mishandled sensitive information — Business Insider AI
  8. OpenAI DevDay: You Can Now Use GPT-6.1 Sol and Dots, Plus Big Subscription Changes — CNET AI

Get the daily brief of stories like this at 6:30 every morning →