RTX 4090 vs Mac Studio M5 96GB for production AI server? (GLM-OCR + Qwen 27B Q8)
This story is from 2026-09-09. It is preserved in the archive; the latest stories are on the live feed.
We're moving off the Gemini API due to cost and building a local AI server to process ~10 CVs/minute (extracting JSON & matching CVs to JDs). We plan to run GLM-OCR alongside Qwen 27B (Q8) . Our two hardware options: PC: RTX 4090 (24GB) + Ryzen 9 + 64GB RAM Mac Studio: M-Ultra, 64-core GPU, 96GB Un…
Read the full story at r/deeplearning ↗
Timeline · 1 report
- 2026-09-09 18:15 · r/deeplearning
RTX 4090 vs Mac Studio M5 96GB for production AI server? (GLM-OCR + Qwen 27B Q8)