Chain-of-Thought Faithfulness: Toggling 'Reasoning Mode' Made One Model 5x More Likely to Follow Its Own Mistakes
This is a submission for the Kaggle Benchmarking Challenge What I Benchmarked A while back I spent time manually poking at Gemini, ChatGPT, and Claude with the same trick: ask a multi-step question, then feed the model a plausible-looking mid-reasoning nudge in the wrong direction and see what happ…
Read the full story at DEV Community — Machine Learning ↗
Timeline · 1 report
- 2026-09-27 07:59 · DEV Community — Machine Learning
Chain-of-Thought Faithfulness: Toggling 'Reasoning Mode' Made One Model 5x More Likely to Follow Its Own Mistakes