AINewsnow

Best small model + setup together with qwen3.8-flash-next

hello, I currently use swift qwen3.8 gsq rco in iq3 xxs in strata. The model is only used for hermes agent. I have 64GB of ram + 32GB R9700 where the model runs on. But I also have an rx7800xt with 16GB. For hermes agent i want to run a small but intelligent enough model as an executioner model, wh…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-10-10 21:54 · r/LocalLLaMA
    Best small model + setup together with qwen3.8-flash-next

More stories

  1. Introducing GPT-6 in ChatGPT with Intelligent UI — OpenAI YouTube
  2. An Anthropic AI model sent a false homicide tip to Philadelphia police — TechCrunch AI
  3. Anthropic bans users from ‘needless abusive or cruel behavior’ towards Claude — The Guardian AI
  4. Philadelphia police receive false homicide tip from Anthropic AI model — The Hill Technology
  5. Qwen Image 2.1 Turbo Released -- Hugging Face — r/StableDiffusion
  6. Impactful scheduling for GPU clusters — Allen Institute for AI (Ai2)
  7. Sophos cuts threat investigation time by 96% with OpenAI Daybreak — OpenAI News
  8. Welcome to Gemini at Work 2026: Introducing the Gemini agent — Google Cloud AI Blog

Get the daily brief of stories like this at 6:30 every morning →