AINewsnow

Built this yesterday with Qwen3.8-Flash-Next (NVFP4, 262K context) on a single NVIDIA DGX Spark

Planning, coding, testing = 8h total. Stack: VSCode Copilot in autopilot mode + SGLang Stats: ∼10k lines generated, ∼800k tokens consumed Sure, it's not GPT-6 Astra level, but for a 100% local ∼180B MoE running on a single DGX Spark at ∼35 tok/s. Not bad...

Read the full story at r/LocalLLaMA ↗

Timeline · 2 reports

  1. 2026-09-19 05:37 · r/LocalLLM
    Has anyone used NVIDIA DGX Spark for serious cybersecurity workloads (Red Team, Blue Team, CTI, GRC)?
  2. 2026-09-18 19:12 · r/LocalLLaMA
    Built this yesterday with Qwen3.8-Flash-Next (NVFP4, 262K context) on a single NVIDIA DGX Spark

More stories

  1. Is ChatGPT currently the best free AI for creating highly realistic images with simple, straightforward prompts? — r/OpenAI
  2. Microsoft director called AI scraping ‘the largest theft of labor in human history,’ while OpenAI head brands ChatGPT an ‘existential threat’ to publishers — revelations come from legal briefs filed in NYT lawsuit — r/artificial
  3. Simulated students that make realistic mistakes help AI tutors learn faster — The Decoder
  4. If I buy the Pro version, will I automatically have access to GPT-6 Astra? — r/ChatGPTPro
  5. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  6. NVIDIA CEO Jensen Huang rejects ‘AI will end the world’ claim, yet cautions ‘we should go as fast as we can but...’ — Mint AI
  7. we made a 27b model for creative writing. performs as good as claude fable 5, at a 40x cheaper price, open weights. — r/GeminiAI
  8. Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI

Get the daily brief of stories like this at 6:30 every morning →