We measured whether our 20 skills actually fire. Baseline recall was 46%, and our first detector only understood Claude's Skill tool.
This story is from 2026-09-10. It is preserved in the archive; the latest stories are on the live feed.
We ship about 20 Claude Code skills. Progressive disclosure means the agent picks one from a single line of description, so an excellent skill nobody picks up is worth zero. We measured ours: micro-recall 46.3%. Over half the prompts that should have fired a skill fired nothing. Not the wrong skill…
Read the full story at r/ClaudeAI ↗
Timeline · 1 report
- 2026-09-10 15:34 · r/ClaudeAI
We measured whether our 20 skills actually fire. Baseline recall was 46%, and our first detector only understood Claude's Skill tool.
More stories
- AI's role in building AI surging? Anthropic says Claude now leads 26% of its R&D — Mint AI
- Security researchers used Claude to help them hack into OpenAI — The Verge AI
- Claude, Anthropic’s AI model, is helping to develop the next version of itself — Fast Company AI
- Is this Minimax H3? — r/StableDiffusion
- OpenAI ‘ethically hacked’ with help of Anthropic’s Claude chatbot — The Guardian AI
- AI skills — r/AI_Agents
- Mathematicians Hate AI. They Can’t Quit It — Wired AI
- Anthropic Shifts Planned IPO to November — Wall Street Journal Technology
Get the daily brief of stories like this at 6:30 every morning →