By overlapping independent computation phases and sharing GPU memory intelligently, you can train vision-language models 1.2–2.2× faster without needing more hardware or changing your RL algorithm.
Rollplex is a GPU runtime that speeds up vision-language model training by overlapping different computational phases.