Senior Software Engineer, RL Post-Training Frameworks
Architect and build scalable reinforcement learning post-training infrastructure that supports the full cycle of training, inference, and rollout across heterogeneous hardware. Work on optimizing distributed systems for performance, fault tolerance, and efficient resource utilization, while contributing to open-source RL frameworks and collaborating with AI researchers and hardware teams to shape future capabilities.