AINewsnow

Dense 27B vs MoE IQ2_XS on one RTX 4090: synthetic bench plus 3 graded real tickets

Rig: RTX 4090 24 GB (450 W), i9-14900K, 32 GB RAM, Linux, desktop on the iGPU. A: Qwen3.8-27B dense on NInfer. 262K context, rk4v4-e8 KV, MTP, 1 slot. B: Qwen3.8-Flash-Next IQ2_XS on Strata 0.1.39. 262K context, q4_0 KV in VRAM, experts in host RAM, MTP. Both models: thinking budget 4096, same prom…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-10-05 03:52 · r/LocalLLM
    Dense 27B vs MoE IQ2_XS on one RTX 4090: synthetic bench plus 3 graded real tickets

More stories

  1. NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI — NVIDIA Blog
  2. Trump’s big AI move: ‘Super Intelligence Force’ launched, Jay Clayton named AI czar — Mint AI
  3. A model guide for the GPT-6 family — OpenAI News
  4. An OpenAI safety employee has quit and is sounding the alarm — The Verge AI
  5. Strata is seriously impressive, running Qwen 3.8 Flash Next on hermes at 512k context. — r/LocalLLM
  6. Introducing Oscilloscope Diffusion — r/comfyui
  7. Apple says it's tightening macOS Full Disk Access' controls due to new risks from AI agents — TechCrunch AI
  8. Google launches satellite to test feasibility of building data centers in space — NPR Technology

Get the daily brief of stories like this at 6:30 every morning →