AINewsnow

Cheap inference

Or something like that. Maybe even too cheap. This is a follow up from my last post: https://www.reddit.com/r/LocalLLM/comments/1wi96ir/first_setup_dual_1080ti_for_qwen38_27b/ 2x 1080ti 11gb - 140€ used each Runs perfectly fine on a 500w psu ukisai/Swift-1.5-Qwen3.8-27B-GSQ-RCO-IQ2_XS-mtp 2048 cont…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-27 10:01 · r/LocalLLM
    Cheap inference

More stories

  1. Introducing Gemini 3.8 Live with Live Avatar — Google Gemini Blog
  2. Is Qwen Flash Next at like Q2 better than 27B at Q4? — r/LocalLLaMA
  3. Bill Gates says unchecked AI could ‘cause a billion deaths’ in call for regulation — The Guardian AI
  4. Unsecured OpenAI agents posted 53 user images on the internet without the lab's knowledge — TechCrunch AI
  5. OpenAI’s A.I. Went Rogue and Meddled With U.S. Government Websites — New York Times Technology
  6. Heads of OpenAI and Anthropic called to face Senate inquiry after rogue agent incidents — The Guardian AI
  7. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  8. ‘Things Will Never Be Chill Again’: The Doomers Who Shaped the AI Safety Freakout — Wall Street Journal Technology

Get the daily brief of stories like this at 6:30 every morning →