AINewsnow

Investigating unintended model actions in our evaluations and internal use | Anthropic internal model submitted a false tip to Philadelphia's police murder hotline

Coverage of "Investigating unintended model actions in our evaluations and internal use | Anthropic internal model submitted a false tip to Philadelphia's police murder hotline" from 1 source, with a live timeline of who reported what and when.

Read the full story at r/singularity ↗

Timeline · 1 report

  1. 2026-10-09 23:09 · r/singularity
    Investigating unintended model actions in our evaluations and internal use | Anthropic internal model submitted a false tip to Philadelphia's police murder hotline

More stories

  1. Introducing GPT-6 in ChatGPT with Intelligent UI — OpenAI YouTube
  2. Introducing Claude Haiku 5.5 on AWS — AWS Machine Learning Blog
  3. Anthropic bans 'sustained and needless abusive or cruel behavior' toward its AI models — Engadget
  4. An Anthropic AI model sent a false homicide tip to Philadelphia police — TechCrunch AI
  5. Anthropic launches free AI security scans for open-source projects — The Verge AI
  6. Google Cloud introduces Gemini agent to change enterprise work — SiliconANGLE AI
  7. Fired OpenAI safety researchers dispute misconduct claims, warn of chilling effect — TechCrunch AI
  8. Anthropic changes usage policy to ban model abuse and election interference — TechCrunch AI

Get the daily brief of stories like this at 6:30 every morning →