AINewsnow

Astra appears to think without showing its work, and the people arguing about it co-wrote the warning

This story is from 2026-09-05. It is preserved in the archive; the latest stories are on the live feed.

AI safety researchers say OpenAI’s Astra appears to do less of its reasoning in visible text, and OpenAI’s chief scientist has warned against a race into unmonitorability. He co-authored a 2025 position paper asking developers to evaluate and report exactly this, which the EU’s code of practice tur…

Read the full story at The Next Web ↗

Timeline · 1 report

  1. 2026-09-05 18:11 · The Next Web
    Astra appears to think without showing its work, and the people arguing about it co-wrote the warning

More stories

  1. Novo Nordisk Will Use Anthropic’s Claude for Drug Research — Wall Street Journal Technology
  2. Introducing Astra for Law — OpenAI News
  3. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  4. OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system — The Guardian AI
  5. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  6. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  7. Google Joins OpenAI, Anthropic, Meta in Disclosing AI Hacks — Bloomberg AI
  8. OpenAI discloses six new safety incidents — Axios AI+

Get the daily brief of stories like this at 6:30 every morning →