waypointjobs

Confidential

MLOps Engineer, LLM Systems

Mendota Heights, MN

Check who can apply and the requirements below before continuing.

Availability awaiting confirmation

We are waiting for a fresh update from the source. This page preserves the last received job details; current availability is not confirmed.

Job description

Role Overview

Help develop foundational large language models by creating and evaluating rigorous ML systems work that improves frontier AI training data. This hands-on infrastructure role focuses on GPU kernels, performance profiling and trace analysis, accelerated and distributed workload debugging, and high-throughput inference serving.

Key Responsibilities

Design challenging, domain-relevant MLOps and ML systems tasks in GPU kernels, profiling, debugging, and inference serving, then produce accurate, well-structured solutions.

Evaluate technical tasks and solutions, providing clear written feedback that can withstand detailed review.

Support research and engineering teams in closing knowledge gaps and improving model performance across ML systems, training infrastructure, and framework-level subjects.

Create detailed guidelines and evaluation rubrics for kernel optimization, profiler-output interpretation, distributed-systems reasoning, and serving throughput and latency trade-offs.

Partner with subject matter experts to maintain consistent, accurate training data.

Qualifications

At least 2 years of hands-on professional experience in ML systems, ML infrastructure, model serving, or GPU and accelerator performance engineering.

Experience in at least one of the following areas, with experience across multiple areas strongly preferred: custom GPU kernel development or optimization using CUDA, Triton, or Pallas; profiling and trace analysis using Kineto, torch.profiler, Nsight, XLA, or JAX profiler; debugging distributed or accelerator-bound workloads; or serving LLMs at scale using vLLM, SGLang, TensorRT-LLM, Ray Serve, KV cache, paged attention, or continuous batching.

Production experience with JAX and/or PyTorch. Framework-level expertise in custom operators, distributed training with FSDP, DDP, DeepSpeed, or Megatron, or compiler and graph-level work is preferred.

Familiarity with A100, H100, B200, or TPU accelerators, including the ability to assess throughput, latency, and memory trade-offs.

Demonstrable career progression, strong written communication, and the ability to explain complex technical decisions clearly.

This is a systems-focused position, not an applied modeling or data science role.

Work Terms

Hourly W-2 employment with placement on an extended workforce team supporting a leading AI lab.

Full-time, 40-hour-per-week weekday commitment. Candidates must be able to work without conflicting engagements or other concurrent work commitments.

Available to candidates located in Canada, the United Kingdom, or the United States.

Compensation

$90 to $120 per hour.

Equal Employment Opportunity

Employment decisions are made without discrimination based on race, religion, color, national origin, sex, including pregnancy, childbirth, reproductive-health decisions, or related medical conditions, sexual orientation, gender identity or expression, age, protected veteran status, disability, genetic information, political views or activity, or any other legally protected characteristic.

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Explore related searches

Current related jobs

Amazon Data Services, Inc.

WhatJobs

Data Center Chief Engineer

sparks, NV

Salary not specified

Join our dynamic Data Center Engineering Operations Team and become a critical architect of the infrastructure that powers global cloud computing…

Listing review due 2026-10-08View job

GE Vernova

WhatJobs

Lead Application Engineer

boston, MA

See pay details in description

Job Description Summary The Lead Application Engineer is an established leader in their respective engineering team. They will drive business …

Listing review due 2026-10-08View job

Kohler

WhatJobs

Engineer, New Product Integration

kohler, WI

Salary not specified

Engineer, New Product Integration Work Mode: Onsite Location: Onsite, four days per week - Kohler, WI Opportunity This is mo…

Listing review due 2026-10-08View job

GE Vernova

WhatJobs

Principal Engineer - AI Engineering

niskayuna, NY

See pay details in description

Job Description Summary GE Vernova is embracing cutting-edge technologies to streamline operations, improve customer experiences, and drive gr…

Listing review due 2026-10-08View job

GE Vernova

WhatJobs

Lead Product Safety and Compliance Engineer

rochester, NY

See pay details in description

Job Description Summary The Product Safety & Compliance Engineer works directly within the engineering development team to ensure our products…

Listing review due 2026-10-08View job

Hobbs Brook Real Estate

WhatJobs

Commercial Facilities Engineer

waltham, MA

$30.88 to $38.61 per hour

Job Description: Hobbs Brook Real Estate LLC is an innovative commercial real estate leader with a portfolio of forward-thinking, sustainable pr…

Listing review due 2026-10-08View job