Does anyone have real experience with Ornith-1.5-9B for coding
This story is from 2026-08-31. It is preserved in the archive; the latest stories are on the live feed.
I'm very happy with Qwen-3.8-27B, I run it on my work machine and switched for major part of real coding tasks from cloud subscriptions to it. I use one dedicated headless RTX 3090, and get maybe 1000-1500 tps prefill, and 45-60 tps generate, with Q8_0 KV and 180224 context. It really beats all clo…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-08-31 12:10 · r/LocalLLaMA
Does anyone have real experience with Ornith-1.5-9B for coding