newsReddit r/MachineLearningTrust 52 · CommunityPublished 11d agoLive · 9d ago
Hyperparameters fine tuning for MARL comparative study [D]
hello everyone. I'm training PPO variants on different multi-agent tasks from the VMAS library (Independent PPO / Graph PPO and such, see HetGPPO by Bettini et al.). I noticed that for every architecture/scenario couple, the optimal hyperparameters sometimes tend to vary (learning rate, entropy coefficient, KL coefficient, SGD batch size, etc). do I need - methodologically s
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 56%optuna/optuna →
- PossiblePossibly related (embedding) · 55%Optim-Agent/optim-agent →
- PossiblePossibly related (embedding) · 49%Let's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Experts →
- PossiblePossibly related (embedding) · 48%algorithmicsuperintelligence/optillm →
- PossiblePossibly related (embedding) · 49%Efficiently Estimating Optimal Hyperparameter Scaling Laws through Power-Law Entropy Search →
- PossiblePossibly related (embedding) · 49%Constrained Hyperparameter Optimization for Streaming Data →
- PossiblePossibly related (embedding) · 60%solegalli/hyperparameter-optimization →
- PossiblePossibly related (embedding) · 50%syne-tune/syne-tune →
Covers
Covers (incoming)
Related across the graph
paperConstrained Hyperparameter Optimization for Streaming Datareposyne-tune/syne-tunerepooptuna/optunarepoOptim-Agent/optim-agentpaperLet's Scale Step by Step: Compute-Efficient Hyperparameter Transfer for Large-Scale Mixture-of-Expertsrepoalgorithmicsuperintelligence/optillmreposolegalli/hyperparameter-optimizationpaperEfficiently Estimating Optimal Hyperparameter Scaling Laws through Power-Law Entropy Search
