About this opportunity
ByteDance lists this Graduate Software Engineer - Cloud-Native GPU Inference opportunity in san jose, California. Review the employer’s description below for duties, qualifications and application requirements.
Job description
ByteDance's Inference Infrastructure team is seeking engineers to design, build, and operate cloud-native GPU-accelerated ML infrastructure at scale, including vLLM, SGLang, and TensorRT-LLM work.
You will join a world-class team within Core Compute Infrastructure, contributing to open-source ecosystems, scheduling, and orchestration across multi-cloud environments. A PhD and deep expertise in distributed systems help you drive high-performance inference at scale.
#J-18808-Ljbffr
Worksite address
san jose, CA, 95199, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.