My local llm when I tell it to do any changes to my vLLM service
better make no mistakes I've been running Qwen 3.8 Flash Next and it's a great driver for Hermes and Pi. I told it to add CUDA_DISABLE_PERF_BOOST=1 to reduce my server's idle power draw
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-22 21:51 · r/LocalLLaMA
My local llm when I tell it to do any changes to my vLLM service