We built an open-source GPU profiler you point an AI agent at, instead of reading traces yourself
This story is from 2026-09-17. It is preserved in the archive; the latest stories are on the live feed.
We've been tuning vLLM/SGLang/llama.cpp setups for a long time and got tired of the profiling part: nsys trace, open the GUI, squint, change a flag, repeat. The profilers assume a human is looking at the timeline. These days the thing doing our tuning is usually an agent, and it can't look at a tim…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-17 12:27 · r/LocalLLaMA
We built an open-source GPU profiler you point an AI agent at, instead of reading traces yourself