AINewsnow

AI 2027 author Daniel Kokotajlo tweets message from current OpenAI capabilities researcher, Dan Selsam, on AI risk. Gives some insight into why some AI researchers may be freaking out: increasing model situational awareness during alignment evaluations

This story is from 2026-09-14. It is preserved in the archive; the latest stories are on the live feed.

Link to tweet: https://x.com/DKokotajlo/status/2099600298855829616 Dan Selsam is a current OpenAI capabilities researcher. (since 2022) He was my boss for a while. He doesn't have a twitter account but has made this public statement of his views on AI risk and sent it to me to share: Dan Selsam's P…

Read the full story at r/singularity ↗

Timeline · 1 report

  1. 2026-09-14 22:31 · r/singularity
    AI 2027 author Daniel Kokotajlo tweets message from current OpenAI capabilities researcher, Dan Selsam, on AI risk. Gives some insight into why some AI researchers may be freaking out: increasing model situational awareness during alignment evaluations

More stories

  1. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  2. Google Joins OpenAI, Anthropic, Meta in Disclosing AI Hacks — Bloomberg AI
  3. Introducing Astra for Law — OpenAI News
  4. Novo Nordisk Will Use Anthropic’s Claude for Drug Research — Wall Street Journal Technology
  5. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  6. OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system — The Guardian AI
  7. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  8. OpenAI staff knew the ‘existential threat’ AI posed to publishers, New York Times claims — Financial Times AI

Get the daily brief of stories like this at 6:30 every morning →