Tag: MARL
-
Multi-Agent RL Guide: Cooperative and Competitive Learning
QMIX vs self-play: benchmarked 3 MARL algorithms on StarCraft micromanagement. One approach dominated cooperative tasks, the other competitive.
-
MARL ์ค์ ๊ฐ์ด๋: QMIX, MAPPO, MADDPG ๊ตฌํ ๋น๊ต
QMIX vs MAPPO vs MADDPG on cooperative tasks: which multi-agent RL algorithm converges faster? Benchmarks from 10M training steps.