You watch what goes into the agent; the data leaves on the way out
This story is from 2026-09-04. It is preserved in the archive; the latest stories are on the live feed.
We ran a two-month internal analysis of agent behavior and found the same attack surface twice: sensitive data leaving on the output side, not the input side. Both incidents followed the same pattern. The agent was behaving normally from an inbound perspective — clean prompts, nothing flagged on th…
Read the full story at r/deeplearning ↗
Timeline · 1 report
- 2026-09-04 18:42 · r/deeplearning
You watch what goes into the agent; the data leaves on the way out