AINewsnow

OpenAI claims GPT-6 Astra is its "most aligned model ever", but OpenAI safety researchers are "very worried Astra is sandbagging/self-sabotaging"

This story is from 2026-09-04. It is preserved in the archive; the latest stories are on the live feed.

Coverage of "OpenAI claims GPT-6 Astra is its "most aligned model ever", but OpenAI safety researchers are "very worried Astra is sandbagging/self-sabotaging"" from 3 sources, with a live timeline of who reported what and when.

Read the full story at r/OpenAI ↗

Timeline · 3 reports

  1. 2026-09-04 17:11 · r/ArtificialInteligence
    OpenAI claims GPT-6 Astra is its "most aligned model ever", but OpenAI safety researchers are "very worried Astra is sandbagging/self-sabotaging"
  2. 2026-09-04 16:59 · r/agi
    OpenAI claims GPT-6 Astra is its "most aligned model ever", but OpenAI safety researchers are "very worried Astra is sandbagging/self-sabotaging"
  3. 2026-09-04 16:57 · r/OpenAI
    OpenAI claims GPT-6 Astra is its "most aligned model ever", but OpenAI safety researchers are "very worried Astra is sandbagging/self-sabotaging"

More stories

  1. Introducing Astra for Law — OpenAI News
  2. OpenAI discloses six new safety incidents — Axios AI+
  3. Microsoft exec called AI scraping the “largest theft of labor in human history” — Ars Technica AI
  4. OpenAI staff knew the ‘existential threat’ AI posed to publishers, New York Times claims — Financial Times AI
  5. How Cooley is accelerating IPO work with ChatGPT — OpenAI News
  6. Helping older adults use AI in everyday life — OpenAI News
  7. How to connect AI usage to business value — OpenAI News
  8. OpenAI launches Astra for Law, a GPT-6 configuration for legal research — SiliconANGLE AI

Get the daily brief of stories like this at 6:30 every morning →