16GB VRAM model test
This story is from 2026-09-09. It is preserved in the archive; the latest stories are on the live feed.
I have given Hermes agent the task to make a test for my local models. Coding and agentic work. The test was done on llama.cpp turboquant fork, all models were run using 131k context. Further optimization of the parameters would still be possible for some of the models. TLDR version: Ornith 1.0 35B…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-09-09 18:48 · r/LocalLLM
16GB VRAM model test