When to Communicate: Belief Distributions and KL Divergence for Principled Gating in Multi-Agent RL
Researchers propose a new approach to communication in multi-agent reinforcement learning. Instead of communicating at every timestep or using a binar...