
vime × RL-Kernel × AMD: Bitwise Train–Rollout Consistency on ROCm
·10 min read
vime and RL-Kernel align selected-token logprobs bit for bit across Megatron training and vLLM rollout on AMD Instinct MI300X, with zero mismatches across 200 GRPO steps.
3 posts

vime and RL-Kernel align selected-token logprobs bit for bit across Megatron training and vLLM rollout on AMD Instinct MI300X, with zero mismatches across 200 GRPO steps.

Announcing ROCm support for vime, now running end-to-end on AMD Instinct MI355X GPUs with prebuilt container.

vime connects slime's training stack with vLLM rollouts to provide a simple, stable, and efficient RL post-training pipeline.