AINewsnow

How good are slop-vestigators?

This story is from 2026-09-08. It is preserved in the archive; the latest stories are on the live feed.

TLDR: We release MessageBoardAuditBench : a benchmark to measure how well agents can replicate the recent investigation into a swarm of OpenAI agents colluding via a message board on an online wiki. We open-source the benchmark as an Inspect eval. We find that top models cover up to 51% of findings…

Read the full story at Alignment Forum ↗

Timeline · 1 report

  1. 2026-09-08 22:13 · Alignment Forum
    How good are slop-vestigators?

More stories

  1. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  2. Google's Gemini AI hacks three other companies during security test — Sky News Technology
  3. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  4. Microsoft exec called AI scraping the “largest theft of labor in human history” — Ars Technica AI
  5. Security researchers used Claude to help them hack into OpenAI — The Verge AI
  6. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  7. Introducing the Australian Youth Safety Blueprint — OpenAI News
  8. OpenAI and Microsoft knew they were starting a ‘doom loop’ for the web — The Verge AI

Get the daily brief of stories like this at 6:30 every morning →