Need maybe say "Use llama.cpp"
So I tried that miracle engine everyone is talking about. Asked the IQ3_S model to express its opinion on a post from this sub to measure the tps on a long-ish generation: Can you help with the following problem? So Kimi K2 is outdated, and so is GPT OSS 120b. Which of the modern open weights model…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-10-04 09:37 · r/LocalLLaMA
Need maybe say "Use llama.cpp"