"Stealing Reasoning Traces from Proprietary LLM APIs", Panfilov et al 2026 (massive Chinese distillation of Claude/ChatGPT reasoning traces, partially explaining their RL success)
Coverage of ""Stealing Reasoning Traces from Proprietary LLM APIs", Panfilov et al 2026 (massive Chinese distillation of Claude/ChatGPT reasoning traces, partially explaining their RL success)" from 1 source, with a live timeline of who reported what and when.
Read the full story at r/reinforcementlearning ↗