papersSEP 10 04:00 UTC
ROTATE: Regret-driven Open-ended Training for Ad Hoc Teamwork
A new arXiv paper presents ROTATE, a training method designed to help agents cooperate effectively with partners they have never seen before, a problem known as ad hoc teamwork. Rather than relying on a pre-built, fixed set of teammate agents followed by a separate coordination stage, the approach uses regret signals to continuously steer the creation of training collaborators. The work addresses generalization in multi-agent reinforcement learning.