Using an LLM to tune the Snake AI's RNN
The LLM is clearly improving the configuration over time. The LLM has proposed over 580 configuration changes, which were executed. The results are fed back to the LLM who further tunes the system. Live Experiment Data (including plots): https://snakeweb.osoyalce.com/ Ax3l Project website: https://…
Read the full story at r/reinforcementlearning ↗
Timeline · 1 report
- 2026-09-20 18:12 · r/reinforcementlearning
Using an LLM to tune the Snake AI's RNN