Research questionHow can decentralized heterogeneous robots combine round-level policy reasoning with tick-level local control without destabilizing navigation learning?Decentralized robots must translate infrequent policy updates into reliable low-level actions while their local controllers continue adapting. Different policy agents and shared feedback add coordination challenges without a central action planner.