AINewsnow

Teaching a 27B Model to Write Trading Alphas: 101 Formulas, 12 Rewards and One Unseen Year

Qwen3.8-27B is trained with multi-reward RL to write one-line trading formulas in the language of WorldQuant’s 101 Formulaic Alphas . A deterministic verifier turns each formula into a daily dollar-neutral long-short book on 46 US stocks and scores it on 12 reward channels. On 50 prompts and a year…

Read the full story at DEV Community — Machine Learning ↗

Timeline · 1 report

  1. 2026-10-10 19:37 · DEV Community — Machine Learning
    Teaching a 27B Model to Write Trading Alphas: 101 Formulas, 12 Rewards and One Unseen Year

More stories

  1. Introducing GPT-6 in ChatGPT with Intelligent UI — OpenAI YouTube
  2. An Anthropic AI model sent a false homicide tip to Philadelphia police — TechCrunch AI
  3. Anthropic bans users from ‘needless abusive or cruel behavior’ towards Claude — The Guardian AI
  4. Philadelphia police receive false homicide tip from Anthropic AI model — The Hill Technology
  5. Anthropic launches free AI security scans for open-source projects — The Verge AI
  6. Impactful scheduling for GPU clusters — Allen Institute for AI (Ai2)
  7. Sophos cuts threat investigation time by 96% with OpenAI Daybreak — OpenAI News
  8. Welcome to Gemini at Work 2026: Introducing the Gemini agent — Google Cloud AI Blog

Get the daily brief of stories like this at 6:30 every morning →