Plimsoll: an agent skill for testing prompt injection, leaks, and tool abuse
This story is from 2026-08-19. It is preserved in the archive; the latest stories are on the live feed.
I’ve been working on LLM/agent security for a while now, mostly around prompt injection, jailbreaks, leaks, tool abuse, and where the actual security boundary sits once a model starts using tools. Getting accepted into Anthropic’s Cyber Verification Program gave me a bit more room to push that work…
Read the full story at r/AI_Agents ↗
Timeline · 1 report
- 2026-08-19 17:42 · r/AI_Agents
Plimsoll: an agent skill for testing prompt injection, leaks, and tool abuse