Qwen3.8-23B-Mini-Me: A Depth-Pruned Qwen3.8-27B (to ~22.7BB)
This story is from 2026-08-19. It is preserved in the archive; the latest stories are on the live feed.
I've been working on a depth pruning approach and decided to try it out on the new Qwen3.8-27B model. I managed to get the model down to about 22.7B params without severe reasoning degradation. No fine-tuning was done, just strategic removal of layers. It's been working well for my use cases in cod…
Read the full story at r/LocalLLaMA ↗
Timeline · 1 report
- 2026-08-19 23:19 · r/LocalLLaMA
Qwen3.8-23B-Mini-Me: A Depth-Pruned Qwen3.8-27B (to ~22.7BB)