SWE-Prime trained on 10% of successful coding trajectories and beat the full set. Are success labels too noisy?
This story is from 2026-08-30. It is preserved in the archive; the latest stories are on the live feed.
SWE-Prime argues that resolved coding-agent runs still contain redundant, ineffective or risky steps. Its two-stage filter selects whole trajectories and then the semantic segments that contribute to learning. The authors report that training on the selected 10% beat the full resolved set on SWE-Be…
Read the full story at r/PromptEngineering ↗
Timeline · 1 report
- 2026-08-30 06:39 · r/PromptEngineering
SWE-Prime trained on 10% of successful coding trajectories and beat the full set. Are success labels too noisy?