Research questionHow can distributed VLA reinforcement learning coordinate variable-latency simulation, inference, and optimization?Synchronous training can leave workers idle when some simulations take longer than others. Variable rollout costs therefore make it difficult to keep simulation, inference, and optimization resources continuously utilized.