AINewsnow

Can current MiniCPM5-2B/any best SOTA under 10B + modern harness beat pre-March 2025 frontier models like Grok 3/GPT-4o?

This story is from 2026-09-16. It is preserved in the archive; the latest stories are on the live feed.

Hey experts! I genuinely have this question and would love a general consensus from people actually using these models. LLMs have advanced a lot in benchmarks and in practical usage. Even GPT-4o and Grok 3 were already enough for general chatting,search lookup, RP, etc. So if you brought those mode…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-16 18:46 · r/LocalLLaMA
    Can current MiniCPM5-2B/any best SOTA under 10B + modern harness beat pre-March 2025 frontier models like Grok 3/GPT-4o?

More stories

  1. The cloud outage that should terrify the CIO — InfoWorld AI
  2. Testing same face with AI models — r/StableDiffusion
  3. chatgpt vs Gemini vs grok (same prompt) — r/GeminiAI
  4. 4-AI relay: ChatGPT drafted, Muse reviewed, Grok and Claude checked the work — every step in the open — r/OpenAI
  5. xAI Ships Grok Voice Transcribe 2.0 With Half the Errors at Same Price — AlphaSignal
  6. ZCode was allegedly caught uploading workspace/.git records to the cloud. — r/LocalLLaMA
  7. AI agents/automation suggestions for a solo biz — r/AI_Agents
  8. Transcription now costs 10 cents an hour. The differentiator is no longer the model, it is whose conversations you get to train on. — The Next Web

Get the daily brief of stories like this at 6:30 every morning →