Research questionHow can reinforcement-learning agents learn robust ad-hoc teamwork without pre-trained partners or hand-tuned partner generation?Agents trained with fixed or costly partner populations may struggle to generalize to unfamiliar teammates. Building diverse training partners can also require manual choices about how those partners are generated.