About this opportunity
GMI Cloud, Inc lists this LLM Inference Optimization Engineer - Frontier Performance opportunity in san francisco, California. Review the employer’s description below for duties, qualifications and application requirements.
Job description
GMI Cloud, Inc is hiring world-class Machine Learning Engineers to advance LLM inference optimization leveraging GPUs and cutting-edge techniques. You will drive research, validation, and productionization of optimization strategies across NV platforms, with a focus on speed, efficiency, and scalability.
You will collaborate with platform and infrastructure teams, contribute to open-source projects, and help define recipes and benchmarks for industry-leading inference performance.
#J-18808-Ljbffr
Worksite address
san francisco, CA, 94199, US
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.