AINewsnow

im making qwen3.8:27b into a MoE and its not half bad...

I have a rather low grade gpu (8 gig laptop GPU) and ive been wanting to run qwen3.8:27b but at the time I simply couldn't. I know that there are some 1-2 bit quants but they dont seem to good. so I made 2 things 1, a custom inference engine made especially for MoE architectures 2, a program that s…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-27 17:22 · r/LocalLLM
    im making qwen3.8:27b into a MoE and its not half bad...

More stories

  1. Scoop: Anthropic's Dario Amodei to have White House dinner with Trump — Axios AI+
  2. Bill Gates says unchecked AI could ‘cause a billion deaths’ in call for regulation — The Guardian AI
  3. Scoop: Top AI companies probing tens of thousands of security incidents — Axios AI+
  4. Unsecured OpenAI agents posted 53 user images on the internet without the lab's knowledge — TechCrunch AI
  5. OpenAI’s A.I. Went Rogue and Meddled With U.S. Government Websites — New York Times Technology
  6. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  7. Is Qwen Flash Next at like Q2 better than 27B at Q4? — r/LocalLLaMA
  8. ‘Things Will Never Be Chill Again’: The Doomers Who Shaped the AI Safety Freakout — Wall Street Journal Technology

Get the daily brief of stories like this at 6:30 every morning →