Small pre-registered experiments testing claims about lesser-known architectures. Most did not hold up
I have been running small, reproducible experiments on architectures that get more attention for their promise than for their measured behaviour. Each tests one specific claim, with the predictions written down in a CRITERIA.md before any results exist and never edited afterwards. Data is either sy…
Read the full story at r/deeplearning ↗
Timeline · 1 report
- 2026-09-26 15:26 · r/deeplearning
Small pre-registered experiments testing claims about lesser-known architectures. Most did not hold up