How we fixed a -0.74 validation inversion and built a zero-shot Pokémon TCG AI (60K params, pure NumPy)
This story is from 2026-09-14. It is preserved in the archive; the latest stories are on the live feed.
Coverage of "How we fixed a -0.74 validation inversion and built a zero-shot Pokémon TCG AI (60K params, pure NumPy)" from 1 source, with a live timeline of who reported what and when.
Read the full story at r/reinforcementlearning ↗
Timeline · 1 report
- 2026-09-14 02:05 · r/reinforcementlearning
How we fixed a -0.74 validation inversion and built a zero-shot Pokémon TCG AI (60K params, pure NumPy)