About this opportunity
NVIDIA lists this RL Post-Training Systems Engineer (Distributed Infra) opportunity in new york, New York. Review the employer’s description below for duties, qualifications and application requirements.
Job description
NVIDIA is building an RL post-training infrastructure team to scale experimental workflows from a single GPU to thousands of nodes, delivering reliable, high-performance runtimes for researchers. You will collaborate with researchers and labs, optimize PyTorch-based RL loops, and improve fault tolerance, elastic scaling, and portability across CPU, GPU, and LPUs.
This role offers the chance to contribute to VeRL, Miles, TorchTitan and related ecosystems while shaping production-grade distributed
#J-18808-Ljbffr
Worksite address
new york, NY, 10261, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.