Ablating 1 of a chess transformer's 128 attention heads causes the model to stop finding the queen sacrifice in a famous chess game. [P]
This story is from 2026-08-23. It is preserved in the archive; the latest stories are on the live feed.
Hooks and reads out Maia-3 23m model with chessformer_lens library: github.com/chessformer-lens/chessformer_lens DOI: 10.5281/zenodo.21986988
Read the full story at r/MachineLearning ↗
Timeline · 1 report
- 2026-08-23 00:22 · r/MachineLearning
Ablating 1 of a chess transformer's 128 attention heads causes the model to stop finding the queen sacrifice in a famous chess game. [P]