This job is closed
Applications are no longer available for this announcement. Explore current related opportunities below.
Job description
Lead the design and implementation of scalable GPU infrastructure and high-performance AI model serving APIs. Architect robust distributed systems to optimize AI workloads and mentor engineers to set the technical vision for AI infrastructure.
Requirements
Requires over 5 years of experience in systems infrastructure with deep expertise in Kubernetes, cloud platforms, and AI serving frameworks. Proficiency in Python, Go, or C++ and a proven track record of managing large-scale GPU fleets are essential.
Key Skills
GPU Infrastructure, AI Model Serving, Distributed Systems, Kubernetes, Cloud-native Infrastructure, API Optimization, Python, Go, C++, GPU Orchestration, TensorFlow Serving, Triton, TorchServe, Model Deployment, MLOps, System Architecture
Benefits
Equity, Comprehensive health benefits, Monthly stipends, Company retreats
#J-18808-Ljbffr
Worksite address
palo alto, CA, 94306, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.