EN #Qwen3-VL-8B 9/27/26 Multimodal Reinforcement Training Process with SkyRL on Amazon SageMaker HyperPod Amazon Web Services demonstrated how to use the HyperPod infrastructure to train the Qwen3-VL-8B model with SkyRL and GRPO. News