About this opportunity
Acceler8 Talent lists this Senior ML Inference Engineer: High-Performance GPU Systems opportunity in northern, Kentucky. Review the employer’s description below for duties, qualifications and application requirements.
Job description
Acceler8 Talent is recruiting an ML Inference Engineer for a Stanford-spun AI startup in San Francisco that is building an eight-figure revenue and growth trajectory. You will design, implement, and optimize the infrastructure powering large-scale LLM workloads and real-time model serving.
Ideal candidates combine Python/C++ proficiency with distributed systems experience, PyTorch expertise, and a passion for low-latency, GPU-accelerated inference at production scale.
#J-18808-Ljbffr
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.