AINewsnow

My own Qwen3.8-27B quants for a 16GB card: 70 t/s at 32K, usable out to 252K

Coverage of "My own Qwen3.8-27B quants for a 16GB card: 70 t/s at 32K, usable out to 252K" from 2 sources, with a live timeline of who reported what and when.

Read the full story at r/huggingface ↗

Timeline · 2 reports

  1. 2026-09-21 12:05 · r/LocalLLM
    My own Qwen3.8-27B quants for a 16GB card: 70 t/s at 32K, usable out to 252K
  2. 2026-09-21 12:05 · r/huggingface
    My own Qwen3.8-27B quants for a 16GB card: 70 t/s at 32K, usable out to 252K

More stories

  1. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  2. Bessent hails US-China AI dialogue ahead of Trump-Xi meeting — Financial Times AI
  3. Anthropic, OpenAI, SpaceXAI, Google sued over call to ‘pace’ AI development — Politico Technology
  4. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  5. Alibaba's open-weight Qwen-Image-2.1 claims to beat closed models in image generation with just 7 billion parameters — The Decoder
  6. Amazon blocks Meta’s Muse AI agent — The Verge AI
  7. NVIDIA CEO Jensen Huang rejects ‘AI will end the world’ claim, yet cautions ‘we should go as fast as we can but...’ — Mint AI
  8. AI hallucination of Chinese nuclear components almost led to US military attack — Ars Technica AI

Get the daily brief of stories like this at 6:30 every morning →