Availability awaiting confirmation
We are waiting for a fresh update from the source. This page preserves the last received job details; current availability is not confirmed.
Job description
NVIDIA’s Local AI team is building the software stack for running large language models and generative AI applications efficiently on NVIDIA edge AI hardware. This role focuses on performance analysis, model validation, and developing inference recipes across multi-node configurations.
The candidate will work with CUDA/C++, Triton, and Python, evaluating new architectures, implementing optimizations, and collaborating with communities and partners to ensure robust model bring-up on NVIDIA GPUs.
#J-18808-Ljbffr
Worksite address
durham, NC, 27703, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.