waypointjobs

The Recruiting Guy

Senior Site Reliability Engineer

washington, DC

Check who can apply and the requirements below before continuing.

About this opportunity

The Recruiting Guy lists this Senior Site Reliability Engineer opportunity in washington, District of Columbia. Review the employer’s description below for duties, qualifications and application requirements.

Job description

Senior Cloud Infrastructure Engineer

Location: San Francisco, CA. Remote unavailable. Modality: On‑Site only. Must live within commuting distance of San Francisco or be willing to relocate. Relocation Assistance: No. Employment Type: Salaried W2 Full‑Time. Salary Range: $175,000 – $250,000.

About the Company

We represent a pioneering open source technology company in San Francisco that is transforming the way creators interact with generative AI.

They are the team behind a powerful, node‑based visual interface that gives artists, developers, and innovators the ability to design, control, and customize AI workflows with complete flexibility.

Their platform allows users to connect modular components, build complex pipelines, and run everything locally with impressive speed and precision.

Their mission is to make generative AI open, transparent, and accessible to everyone. Built around community collaboration and creative empowerment, their tools help users experiment freely and bring their ideas to life.

Whether it is visual storytelling, image generation, or advanced machine learning, their technology gives creators the freedom to explore without limitations.

About the Role

In this role, you will take the lead on designing, deploying, and maintaining large‑scale distributed systems that power AI workloads. You will work closely with core engineers to shape the company’s long‑term infrastructure vision while ensuring scalability, performance, and reliability across environments.

What You’ll Do

Design, build, and maintain the core infrastructure that powers AI workloads at scale

Manage and automate GPU compute clusters using tools such as Python, Kubernetes, Terraform, and Ansible

Architect and operate systems for orchestration, observability, distributed storage, and networking

Ensure reliability, scalability, and performance across production environments

Collaborate closely with core engineers to design infrastructure for new features and systems

Contribute to technical strategy and long‑term infrastructure vision

Drive best practices for infrastructure automation, deployment, and monitoring

Requirements

5+ years of experience as an Infrastructure Engineer or Site Reliability Engineer building and operating large‑scale distributed systems

Skilled in Python and comfortable working with infrastructure‑as‑code tools such as Terraform and Ansible

Familiar with container orchestration systems such as Kubernetes and related tooling like FluxCD, Prometheus, and Grafana

Capable of managing high‑performance GPU environments across cloud and bare metal setups

Highly adaptable, resourceful, and motivated by building things from the ground up

Excited to work in a small, fast‑growing team where autonomy and accountability are key

Comfortable working on‑site in a startup setting where collaboration and speed matter most

Bonus Points

Experience contributing to or maintaining open‑source projects

Background working with AI infrastructure, ML pipelines, or GPU orchestration

Strong computer science fundamentals and ability to work across different programming languages or frameworks

Skills

fluxcd, ansible, kubernetes, grafana, prometheus, python, terraform, infrastructure

Seniority Level

Mid‑Senior level

Employment Type

Full‑time

Job Function

Engineering and Information Technology

Industries

Human Resources Services

#J-18808-Ljbffr

Worksite address

washington, DC, 20022, US

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Ready for your next step?Apply on the official website
Apply on WhatJobs ↗

Explore related searches

Current related jobs

gpac

WhatJobs

Drywall Project Engineer: $75K-$90K

sparks, NV

See pay details in description

Job Description JOB DESCRIPTION: $75K-$90K SEEKING COMMERCIAL DRYWALL PROJECT ENGINEERS GPAC: #1 Commercial Drywall Recruiting Firm in Nort…

Last received from source 2026-10-06View job

Cushman & Wakefield

WhatJobs

UNION Operating Engineer

philadelphia, PA

See pay details in description

Job Title UNION Operating Engineer Job Description Summary Responsible to ensure the efficient operation and maintenance of mechanical, e…

Last received from source 2026-10-06View job

SmartRecruiters, Inc.

WhatJobs

Senior Security Solutions Engineer (Pre-Sales)

st. louis, MO

Salary not specified

Check Point is seeking a seasoned pre-sales engineer to partner with the sales team in identifying prospects and showcasing our comprehensive sec…

Last received from source 2026-10-06View job

SmartRecruiters, Inc.

WhatJobs

Security Engineer

st. louis, MO

Salary not specified

At Check Point, what you do matters. Every day, we protect over 100,000 organizations worldwide from increasingly sophisticated cyber and AI-driv…

Last received from source 2026-10-06View job

JCPenney

WhatJobs

Senior Engineer -Creative Marketing

dallas, TX

$147,750.00 /Yr

Senior Engineer, Creative Marketing Role Overview Catalyst Brands is seeking a Senior Engineer, Creative Marketing to help us build platform …

Last received from source 2026-10-06View job