AINewsnow

For those of you forced to only use open models from Western labs in production, what are you deploying?

This story is from 2026-09-12. It is preserved in the archive; the latest stories are on the live feed.

First off, I know that GLM, Qwen, and DeepSeek absolutely dominate in terms of SOTA Open Source models, and that’s what I use in my personal projects and for school, however, I’m also responsible for deploying local AI on my organization’s H100s, and we are forbidden by management from running any…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-12 16:17 · r/LocalLLaMA
    For those of you forced to only use open models from Western labs in production, what are you deploying?

More stories

  1. A company ran 8 identical AI societies for weeks with different models and just published what happened. Some of it is genuinely unsettling. — r/ArtificialInteligence
  2. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  3. US government website used Chinese model the FBI called "malicious" — Ars Technica AI
  4. Cactus Needle 3: A Sliceable 8-29MB Automation Foundation Model That Matches DeepSeek v4 Flash — r/LocalLLaMA
  5. Qwen 3.8 27B Running for 63 hours on a RTX 3090 to solve the Riemann hypothesis — r/LocalLLM
  6. Deployed Qwen 3.6 35B A3B on a single DGX Spark supporting 12 concurrent users at 262K context. Are there better ways to optimize this? — r/LocalLLM
  7. Qwen Developers on X: "Qwen-Image 2.1 is going open source" — r/StableDiffusion
  8. Qwen q4 3.8 27b 16 tok/s 32k RTX 3060 :D — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →