About this opportunity
uRun lists this Founding ML Inference Performance Engineer opportunity in san francisco, California. Review the employer’s description below for duties, qualifications and application requirements.
Job description
uRun, located in San Francisco, is seeking a founding ML Performance Engineer to drive AI infrastructure performance. In this role, you will write custom CUDA kernels and optimize model inference for real-time applications, significantly impacting performance across the stack.
The ideal candidate will possess deep knowledge of CUDA, experience with AI workloads, and a strong capacity for optimization. The position offers a competitive salary, equity, and top-tier tools for an exceptional contributor.
#J-18808-Ljbffr
Worksite address
san francisco, CA, 94199, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.