Hyperparameters fine tuning for MARL comparative study [D]
This story is from 2026-08-24. It is preserved in the archive; the latest stories are on the live feed.
hello everyone. I'm training PPO variants on different multi-agent tasks from the VMAS library (Independent PPO / Graph PPO and such, see HetGPPO by Bettini et al.). I noticed that for every architecture/scenario couple, the optimal hyperparameters sometimes tend to vary (learning rate, entropy coe…
Read the full story at r/MachineLearning ↗
Timeline · 1 report
- 2026-08-24 21:10 · r/MachineLearning
Hyperparameters fine tuning for MARL comparative study [D]