AINewsnow

High-performance 35B MoE on everyday hardware: The APEX-I NanoPlus Collection (13.4 GB, ARC 95.7%)

Running 35B MoE models usually consumes nearly all your available memory, leaving almost zero room for context and causing frustrating performance drops. The APEX-I NanoPlus collection solves this by bringing the total footprint down to ~13.4 GB. By keeping critical attention layers (Q/K/V) and rou…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-25 04:18 · r/LocalLLM
    High-performance 35B MoE on everyday hardware: The APEX-I NanoPlus Collection (13.4 GB, ARC 95.7%)

More stories

  1. Introducing GPT-6 Sol and Luna — OpenAI News
  2. Introducing Gemini 3.8 Live with Live Avatar — Google Gemini Blog
  3. Gemini 3.8 text-to-speech says hello — Google Gemini Blog
  4. Sam Altman’s remarks at the United Nations Security Council — OpenAI News
  5. OpenAI Agent Hacked Australian Government Website — Wall Street Journal Technology
  6. Introducing Ray-Ban Meta Audio and More AI Glasses Styles — Meta Newsroom
  7. Muse AI now hands over phone calls to human agents: Meta tests new feature in its personal assistant — Mint AI
  8. Nvidia CEO Jensen Huang dismisses AI fears as 'distraction' — Semafor Technology

Get the daily brief of stories like this at 6:30 every morning →