Ollama runs the model. Who runs the companion?
This story is from 2026-09-12. It is preserved in the archive; the latest stories are on the live feed.
Most of this sub has local inference figured out. Ollama, llama.cpp, vLLM, whatever fits the VRAM. The gap I keep hitting: every session still starts colder than it should. Chat history ≠ companion memory. RAG on docs ≠ remembering that last month’s fix for the homelab broke DNS. Chatbots reset. Co…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-12 18:42 · r/LocalLLM
Ollama runs the model. Who runs the companion?