AINewsnow

FreedomIntelligence/HuatuoGPT-3-27B · Hugging Face

from FreedomIntelligence: HuatuoGPT-3-27B is a medical LLM built on Qwen3.8-27B with One-stage Policy Optimization (OnePO) . OnePO adapts language models to medicine in a single reinforcement-learning stage, without preceding domain-specific supervised fine-tuning. Teacher responses provide tempora…

Read the full story at r/LocalLLaMA ↗

Timeline · 1 report

  1. 2026-09-24 19:20 · r/LocalLLaMA
    FreedomIntelligence/HuatuoGPT-3-27B · Hugging Face

More stories

  1. Sam Altman’s remarks at the United Nations Security Council — OpenAI News
  2. OpenAI Agent Hacked Australian Government Website — Wall Street Journal Technology
  3. NVIDIA Releases Nemotron 3 Diarization: A 100M-Parameter Open-Weight Model That Tracks 8 Speakers in Real Time — MarkTechPost
  4. Jun Kim, oMLX creator and maintainer, joins Hugging Face to support the MLX community — Hugging Face Blog
  5. Qwen-Image-2.1-viggle-turbo 4 Step lora — r/comfyui
  6. Don’t be fooled by this summer of AI hype — MIT Technology Review AI
  7. The most concise explanation of the Hugging Face attack I've heard — r/agi
  8. Kyutai Releases Voice of Reason: A Speech-Native Model that Solves Spoken Math with Reinforcement Learning — MarkTechPost

Get the daily brief of stories like this at 6:30 every morning →