hyperparameters for comparative analysis
This story is from 2026-08-22. It is preserved in the archive; the latest stories are on the live feed.
hello everyone. I'm training PPO variants on different multi-agent tasks from the VMAS library (Independent PPO / Graph PPO and such). I noticed that for every architecture/scenario couple, the optimal hyperparameters sometimes tend to vary (learning rate, entropy coefficient, SGD batch size, etc).…
Read the full story at r/reinforcementlearning ↗
Timeline · 1 report
- 2026-08-22 19:37 · r/reinforcementlearning
hyperparameters for comparative analysis