AINewsnow

One epoch of domain fine-tuning took Qwen2.5-Coder-14B from 1% to 92% compile success on MQL5. gpt-5.6-sol got 97% on the same items. Benchmark is public.

This story is from 2026-09-05. It is preserved in the archive; the latest stories are on the live feed.

I build MQL5 training data (the MetaTrader 5 trading language —niche, thin public corpus, easy to get subtly wrong). For the last 11 months I've been generating specs with my own generators and machine-verifying every completion through the real compiler and a back test pipeline. Today I published…

Read the full story at r/LocalLLM ↗

Timeline · 1 report

  1. 2026-09-05 08:08 · r/LocalLLM
    One epoch of domain fine-tuning took Qwen2.5-Coder-14B from 1% to 92% compile success on MQL5. gpt-5.6-sol got 97% on the same items. Benchmark is public.

More stories

  1. Meet the Data Agent in ChatGPT Work — OpenAI YouTube
  2. Microsoft and OpenAI Workers Worry About ‘Largest Theft of Labor’ in History — New York Times Technology
  3. we made a 27b model for creative writing. performs as good as claude fable 5, at a 40x cheaper price, open weights. — r/GeminiAI
  4. Anthropic mulls new AI model ahead of IPO to counter OpenAI's GPT-6 Astra, says report: What we know — Mint AI
  5. what's the state of the art recipe for running Qwen3.8-Flash-Next with a pair of 3090s and a ton of system RAM rn? — r/LocalLLaMA
  6. Gemini 4 Pro vs Fable 5 vs GPT6 Astra — r/GeminiAI
  7. I ran Claude code and Codex in parallel for 15 days. Here's what I found. — r/AI_Agents
  8. ChatGPT for Word is now available — OpenAI YouTube

Get the daily brief of stories like this at 6:30 every morning →