qwen4exp: add hc ops by am17an · Pull Request #28901 · ggml-org/llama.cpp
This story is from 2026-09-16. It is preserved in the archive; the latest stories are on the live feed.
time to re-benchmark Qwen Flash Next again
Read the full story at r/LocalLLaMA ↗
Timeline · 3 reports
- 2026-09-16 18:36 · r/LocalLLaMA
Enable CUDA graph for MTP draft by gaugarg-nv · Pull Request #28549 · ggml-org/llama.cpp - 2026-09-16 15:51 · r/LocalLLaMA
Release b11003 · ggml-org/llama.cpp - 2026-09-16 09:51 · r/LocalLLaMA
qwen4exp: add hc ops by am17an · Pull Request #28901 · ggml-org/llama.cpp