I built an open-source eval + monitoring tool
LitmusAI tests AI agents before you deploy them and monitors them while they run. Write test cases in YAML or Python. It checks correctness, consistency across runs, latency, and token cost, and plugs into CI. The runtime monitor watches live traffic for prompt injection and policy violations and s…
Read the full story at r/AI_Agents ↗
Timeline · 1 report
- 2026-10-02 06:29 · r/AI_Agents
I built an open-source eval + monitoring tool