waypointjobs

SmartRecruiters, Inc.

Machine Learning Platform Engineer, Machine Learning (ML) and Artificial Intelligence (AI) Required, Work From Home

workfromhome, CA

Check who can apply and the requirements below before continuing.

About this opportunity

SmartRecruiters, Inc. lists this Machine Learning Platform Engineer, Machine Learning (ML) and Artificial Intelligence (AI) Required, Work From Home opportunity in workfromhome, California. Review the employer’s description below for duties, qualifications and application requirements.

Job description

Machine Learning Platform Engineer, Machine Learning (ML) and Artificial Intelligence (AI) Required, Work From Home

Full-time

Compensation: USD 140,000 - USD 180,000 - yearly

Machine Learning Platform Engineer, Machine Learning (ML) and Artificial Intelligence (AI) Required, Work From Home

As the Machine Learning Platform Engineer, you will build the infrastructure and systems that power Artificial Intelligence (AI) capabilities. You will design and operate the systems behind the Artificial Intelligence (AI) stack, from model training and evaluation to deployment, inference, observability, and continuous improvement. You will work closely with Artificial Intelligence (AI) engineers, researchers, and product engineers to turn models into reliable, scalable, and cost-efficient production systems. You will build the platforms, tooling, and infrastructure that enable the team to experiment quickly and bring AI capabilities to production with confidence. Machine Learning (ML) and Artificial Intelligence (AI) experience are required. This position is 100% Remote.

MUST BE WILLING TO TAKE A 60 MINUTE CODING ASSESSMENT.

- Build and operate the Machine Learning (ML) infrastructure and platforms powering Artificial Intelligence (AI) products.

- Design systems for model training, evaluation, deployment, inference, and experimentation.

- Build and optimize model serving and inference infrastructure for high-throughput and low-latency workloads.

- Improve reliability, scalability, latency, and cost efficiency of Artificial Intelligence (AI) systems.

- Develop reliable pipelines for data preparation, training, evaluation, model release, and continuous improvement.

- Build platforms and tooling that enable Artificial Intelligence (AI) engineers and researchers to experiment, evaluate, and ship models faster.

- Develop evaluation and benchmarking infrastructure to measure model quality, performance, and regressions.

- Build production observability, monitoring, tracing, and alerting for Artificial Intelligence (AI)/Machine Learning (ML) workloads.

- Improve Artificial Intelligence (AI) systems across reliability, scalability, latency, throughput, and cost.

- Identify bottlenecks across the Machine Learning (ML) stack and continuously improve system performance.

- Work closely with Artificial Intelligence (AI) engineers, researchers, and product teams to turn evolving model requirements into production-ready infrastructure.

- AI infrastructure reliably supports production workloads at scale.

- Models can be trained, evaluated, deployed, and improved efficiently.

- Inference systems deliver strong latency, throughput, reliability, and cost efficiency.

- Machine Learning (ML) pipelines are reproducible, observable, maintainable, and robust.

- Model and infrastructure regressions are detected quickly and diagnosed efficiently.

- Common Machine Learning (ML) infrastructure capabilities become reusable platform primitives rather than being rebuilt for every AI product.

- The AI stack can evolve rapidly as new models, architectures, and inference techniques emerge.

Machine Learning Platform Engineer Qualifications:

- Machine Learning (ML) and Artificial Intelligence (AI) experience are required.

- Strong software engineering fundamentals and experience building production systems.

- Experience building Machine Learning (ML) infrastructure, platforms, or production machine learning systems.

- Experience with model deployment, inference, evaluation, or data pipelines.

- Strong understanding of distributed systems and system reliability.

- Ability to write clean, maintainable, production-quality code.

- Comfortable working in ambiguous, fast-moving environments.

- Bias toward ownership, experimentation, and continuous improvement.

- Tech Stack: Python, PyTorch, JAX, LLM and ML serving infrastructure such as vLLM, SGLang, or TensorRT-LLM, Cloud infrastructure, Distributed systems, Machine Learning (ML)/data pipelines and workflow orchestration, GPU infrastructure and performance tooling, and Vector databases and retrieval infrastructure.

Benefits include medical insurance, Dental, Vision, Savings Plan Options, PTO, etc.

Keywords: San Francisco CA Jobs, AI, Artificial Intelligence, Cloud Infrastructure, Data Pipelines, Distributed Systems, GPU Infrastructure, JAX, LLM, Large Language Model, Machine Learning Platform Engineer, ML, Machine Learning, Python, PyTorch, SGLang, TensorRT-LLM, Vector Databases, vLLM, Workflow Orchestration, Work From Home, Remote, California Recruiters, IT Jobs, California Recruiting

All your information will be kept confidential according to EEO guidelines.

#J-18808-Ljbffr

Worksite address

workfromhome, CA, 94199, US

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Ready for your next step?Apply on the official website
Apply on WhatJobs ↗

Explore related searches

Current related jobs

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Space Systems Mechanical Engineer

laurel, MD

See pay details in description

Description Do you want to design and build unique space structures and spacecraft for NASA missions that enable groundbreaking scientific disco…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

System Security Engineer

laurel, MD

See pay details in description

Description Are you looking for an opportunity to utilize your technical skills to solve complex, real-world problems? If so, we're looking …

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Advanced Reentry Mission Engineer

laurel, MD

See pay details in description

Description Are you interested in hypersonic and reentry system design and prototyping? Do you want to make contributions to next generation …

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Thermal and EO/IR Modeling and Simulation Engineer

laurel, MD

See pay details in description

Description Are you looking for a unique opportunity to impact significant advances to the nation's groundbreaking integrated air and missile de…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Network Effects Engineer

laurel, MD

See pay details in description

Description Do you want to perform advanced research, development, and test & evaluation of communications systems and network technologies that…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

System Realization and Resilience Engineer

laurel, MD

See pay details in description

Description Are you passionate about applying system engineering principles to influence the development and resilience of future strategic weap…

Listing review due 2026-10-06View job