ATraining's Rocket asynchronous RL infrastructure — controller, problem workers, rollout workers, router + SGLang inference, and the weight-transfer plan between differently-sharded learner and inference fleets. Use when architecting large-scale async RL systems, balancing inference vs learner GPUs, stabilizing inference at scale, or closing the train/inference numerical gap.