AINewsnow

I fine tuned Gemma 4 12B for a 2.7x improvement on tool calling because I can't fit anything else comfortably into my 16 GBs of Vram

This story is from 2026-08-23. It is preserved in the archive; the latest stories are on the live feed.

Gemma 12B is obviously a very well trained model, I always thought the fine tuning they did on it wasn't really cut out for agentic coding. From my own experiences it struggles to use the tools it's given from Github Copilot and is also very inept at the cli too. So I thought I'd kill two birds wit…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-08-23 01:30 · r/LocalLLaMA
    I fine tuned Gemma 4 12B for a 2.7x improvement on tool calling because I can't fit anything else comfortably into my 16 GBs of Vram

More stories

  1. OpenAI and Microsoft knew they were starting a ‘doom loop’ for the web — The Verge AI
  2. OpenAI's latest AI revelation is a 'serious situation,' Microsoft's Suleyman tells CNBC — CNBC Technology
  3. Microsoft director called AI scraping ‘the largest theft of labor in human history,’ while OpenAI head brands ChatGPT an ‘existential threat’ to publishers — revelations come from legal briefs filed in NYT lawsuit — Tom's Hardware
  4. Running Qwen3.8-Flash-Next ~85GB GGUF on 2× RTX 3060 12GB: ~12 tok/s, 131k ctx, CPU MoE, and a 26.5k agent prompt — r/LocalLLM
  5. Plugin4Shell and NIST IR 8587, days apart: what actually authorizes an AI agent’s action? — r/AI_Agents
  6. Microsoft AI Chief Says China Isn’t Excuse to Forego Regulation — Bloomberg AI
  7. AI Firms Knew Chatbots Were an ‘Existential Threat’ to Journalists, Court Docs Show — CNET AI
  8. A zero-click RCE flaw in AI coding agents could have exposed enterprise systems — InfoWorld AI

Get the daily brief of stories like this at 6:30 every morning →