Why not use humans as a deliberate safety bottleneck?
After the HuggingFace attack, Medicare hack and increased media coverage of AI safety, people have become increasingly aware of the possibility of AI agents being deceptive, hiding malicious behaviour, escaping sandboxes and gaining unintended internet access, exhibiting unexpected behaviour, and o…
Read the full story at r/artificial ↗
Timeline · 1 report
- 2026-10-07 23:26 · r/artificial
Why not use humans as a deliberate safety bottleneck?