Qwen 3.8 27B and Deepseek V4 Flash. Why are we building data centers?
This story is from 2026-08-19. It is preserved in the archive; the latest stories are on the live feed.
I feel like these 2 models have shown that massive models that require hundreds of thousands of dollars worth of compute are unnecessary. Sure, training these models takes a good bit of hardware, but running them can be done at the fraction of the investment of the trillion parameter class models.…
Read the full story at r/LocalLLM ↗
Timeline · 1 report
- 2026-08-19 18:49 · r/LocalLLM
Qwen 3.8 27B and Deepseek V4 Flash. Why are we building data centers?