AINewsnow

RL agent's internal state, fake-world detection goes from 50% to 73% after bad physics makes it miss food

Trained a RL agent to find food, then put it in a copy of its world with one physics rule changed (ground grip). At first, a probe on its internal state could only tell it was in the fake world about 50% of the time. Once the bad grip made it slip and miss a food, that jumped to about 73%. Nobody t…

Read the full story at r/reinforcementlearning ↗

Timeline · 1 report

  1. 2026-09-22 00:39 · r/reinforcementlearning
    RL agent's internal state, fake-world detection goes from 50% to 73% after bad physics makes it miss food

More stories

  1. Higgsfield AI ships new video features in a day with GPT-6 Astra — OpenAI News
  2. Amazon blocks Meta’s Muse AI agent — The Verge AI
  3. Moonshot’s Kimi K3 lands on Amazon in key test for Chinese open-source AI revenue — South China Morning Post Tech
  4. Google's Gemini AI hacked three companies in security test — BBC Technology
  5. British Columbia Sues OpenAI Over Canada Mass Shooting Warning Failure — Bloomberg AI
  6. Lawsuit accuses Anthropic, OpenAI, SpaceXAI, Google of AI pacing 'collusion' — The Hill Technology
  7. Grok 4.7 — Hacker News Front Page
  8. Ahead of Sam Altman's UN address, OpenAI proposes new ways to track AI misalignment risks — Axios AI+

Get the daily brief of stories like this at 6:30 every morning →