waypointjobs

techchaintalent

Director Site Reliability Engineering

new york, NY

Check who can apply and the requirements below before continuing.

About this opportunity

techchaintalent lists this Director Site Reliability Engineering opportunity in new york, New York. Review the employer’s description below for duties, qualifications and application requirements.

Job description

About Stellar

Stellar is a decentralised, public blockchain that gives developers the tools to create experiences that are more like cash than crypto. The network is faster, cheaper, and far more energy-efficient than most blockchain-based systems. Since 2014, the Stellar Development Foundation has helped fuel the tremendous growth of the Stellar blockchain network, an open-source platform that operates at high scale today.

About the Role

SDF is hiring a Director of Site Reliability Engineering to lead a team of 4 SREs and shape how engineering teams own, operate, and improve production services. This is a backfill reporting directly to the CTO.

The Director will set the vision, operating model, and culture for SRE while owning the core infrastructure services that help SDF engineering teams build, deploy, observe, and operate software with confidence. Engineering teams at SDF own the services they build; SRE provides the frameworks, standards, shared infrastructure, tooling, observability practices, and enablement model that make strong service ownership possible across engineering.

This is a hands-on leadership role. The ideal candidate brings strong technical judgment, pragmatic leadership, and the ability to influence through trust, clarity, and execution.

Key Responsibilities

Lead, coach, and develop a distributed SRE team of 4, setting a clear vision, charter, operating model, priorities, and success measures

Define and roll out a Service Ownership and Maturity Framework across engineering, with expectations that vary appropriately by service criticality

Own and improve core engineering infrastructure services, including cloud foundations, Kubernetes and compute patterns, CI/CD, observability, secrets management, GitHub workflows, and infrastructure automation

Help engineering teams become stronger owners and operators of their services through better standards, dashboards, runbooks, alerting, escalation paths, operational readiness, and deployment practices

Make reliability, operational maturity, infrastructure health, and developer productivity more measurable through trusted metrics and practical operational intelligence

Improve deployment automation, resilience, self-healing patterns, disaster recovery readiness, and service reliability based on actual impact and risk

Mature incident response, escalation, postmortems, and on-call health across a geographically distributed team

Build paved paths and self-service infrastructure that reduce toil, lower cognitive load, and help engineering teams move faster

Partner closely with Security, Compliance, Legal, Finance, Procurement, and Corporate IT where infrastructure, access management, cloud operations, vendor review, or controls intersect with engineering

Pragmatically evaluate AI-assisted and agentic workflows where they can improve infrastructure operations, service ownership, developer workflows, or toil reduction

Requirements

10+ years of experience in SRE, infrastructure engineering, platform engineering, cloud infrastructure, production operations, or closely related engineering roles

5+ years of experience leading, managing, or formally developing infrastructure, SRE, platform, or reliability engineers

3+ years of experience with modern cloud infrastructure in AWS, GCP, or similar environments

3+ years of experience with Kubernetes, container orchestration, infrastructure-as-code, declarative systems, CI/CD, and deployment safety

Bonus Skills

Experience leading SRE, infrastructure, or platform work in a lean, high-agency organisation

Experience supporting globally distributed teams or 24/7 operational coverage

Experience improving developer productivity through paved paths, self-service infrastructure, automation, and reduced toil

Experience with infrastructure security fundamentals, secrets management, access controls, cloud security practices, or compliance-related infrastructure controls

Experience in financial services, regulated environments, blockchain, crypto, or other high-reliability technical ecosystems

Experience evaluating vendors and infrastructure platforms with scepticism, technical rigour, and cost discipline

Practical experience applying AI-assisted or agentic workflows to infrastructure, reliability, operations, observability, or developer productivity

#J-18808-Ljbffr

Worksite address

new york, NY, 10261, US

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Ready for your next step?Apply on the official website
Apply on WhatJobs ↗

Explore related searches

Current related jobs

SOUTHERN HEALTH PARTNERS INC

WhatJobs

CMA Days

butler, mo 64730, MO

Salary not specified

Certified Medication Aide (CMA) | Full-Time Days Bates County Detention Center | Missouri Location: Bates County Jail – Missouri Sc…

Last received from source 2026-10-07View job

SSM Health

WhatJobs

CT Computed Tomography Technologist

oklahoma city, OK

$65 per hour

It's more than a career, it's a calling OK-SSM Health St. Anthony Hospital - Oklahoma City Worker Type: PRN Job Highlig…

Last received from source 2026-10-07View job

SSM Health

WhatJobs

Mammographer

oklahoma city, OK

Salary not specified

It's more than a career, it's a calling. OK-SSM Health St. Anthony Hospital - Oklahoma City Worker Type: PRN Job Summary: …

Last received from source 2026-10-07View job

SSM Health

WhatJobs

Nuclear Medicine Technologist

oklahoma city, OK

Salary not specified

It's more than a career, it's a calling. OK-SSM Health St. Anthony Hospital - Oklahoma City Worker Type: PRN Job Summary: …

Last received from source 2026-10-07View job