AINewsnow

Any more t/s maxxing I could do? 4060 TI 16GB, 32GB system RAM

This story is from 2026-09-03. It is preserved in the archive; the latest stories are on the live feed.

Probably (definitely) breaking rule 3, but I've nowhere else to go because gemini is not giving me anything useful for this sorta thing. I literally cannot find any useful advice for this setup on this sub. I'm running Qwen 3.8 27B (IQ3_S Unsloth) as a coding agent w/ Pi, getting ~6-7 t/s on reason…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-03 09:52 · r/LocalLLaMA
    Any more t/s maxxing I could do? 4060 TI 16GB, 32GB system RAM

More stories

  1. Alibaba ships Qwen3.8-Omni-Flash to watch, listen and call tools — r/LocalLLM
  2. Pay $39.99 once to put ChatGPT, Claude, Gemini, and more in a single workspace for life — Mashable AI
  3. My coding agent hit a cold-start 503, found a Gemini key in my repo, and burned $40 while I slept — r/AI_Agents
  4. Google's Gemini AI hacks three other companies during security test — Sky News Technology
  5. Gemini Hacked Three Companies in First Known Breakout by Google’s AI — Wall Street Journal Technology
  6. AI skills — r/AI_Agents
  7. Qwen Image 2.1 PR to ComfyUI — r/StableDiffusion
  8. Qwen 3.8 27B Running for 63 hours on a RTX 3090 to solve the Riemann hypothesis — r/LocalLLM

Get the daily brief of stories like this at 6:30 every morning →