Can we get some quality control on all these model perf posts?
Too many hyperactive amateurs are coming in here with "1b model at 832843tok/s!" and hardly any of them have all the info necessary for local runners to evaluate. We need context ladders with perplexity/KLD, hardware specs, model params and quant(s), runtime, tuned runtime parameters, basically eve…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-09-28 20:32 · r/LocalLLaMA
Can we get some quality control on all these model perf posts?