waypointjobs

Specter

Site Reliability Engineer

san francisco, CA

Check who can apply and the requirements below before continuing.

About this opportunity

Specter lists this Site Reliability Engineer opportunity in san francisco, California. Review the employer’s description below for duties, qualifications and application requirements.

Job description

Company Background

Specter's mission is to help automate the physical world.

Today, we build video sensors with state‑of‑the‑art AI agents that answer any question, anywhere in their environments. Our systems can automatically detect and reason about any physical activity captured on camera, from security incidents (e.g. perimeter intrusion, theft, LPR), to safety monitoring (e.g. PPE detection, injured people), to operational efficiency (e.g. material tracking, congestion monitoring). We offer both long‑range wireless (1km range) and wired sensor variants to suit any deployment.

Our co‑founders Xerxes and Philip are passionate about empowering our partners in the fast approaching world of physical AI and robotics. We are a small, fast growing team who hail from Anduril, Tesla, Uber, and the U.S. Special Forces.

The Role

We're hiring a Site Reliability Engineer to own the operational health of our connected sensor platform — spanning a live fleet of edge hardware deployed at customer sites and the cloud infrastructure behind it.

This is a high‑ownership role at the intersection of ops and platform engineering. You'll drive reliability across our sensor fleet — triaging issues in the field, building the systems that prevent them from recurring, and owning the observability that keeps us ahead of problems as we scale.

You set your own priorities across all three:

Responsibilities

Reactive — Triage & Recovery

Debug and triage issues across a live fleet of diverse Linux‑based sensor nodes and edge appliances deployed at customer sites.

SSH into field hardware to diagnose, patch, and recover systems — often with limited remote access and incomplete information.

Own site bring‑ups end to end; be the person who gets things back online.

Systems Builder — Close the Loop

Build and maintain fleet management systems: OTA update pipelines, device health tracking, remote diagnostics, and lifecycle tooling.

Identify repeat fires and eliminate them — build tooling, pre‑deployment checks, and root cause processes that prevent recurrence.

Automate toil relentlessly: if you're doing something twice, you should be scripting it.

Collaborate with embedded systems, and platform teams to define reliability and deployment requirements.

Observability Owner — Fleet Visibility

Design and implement observability (logging, metrics, alerting) across edge devices and cloud infrastructure (AWS).

Surface and close telemetry gaps; build fleet‑wide visibility that enables data‑driven reliability decisions.

Develop runbooks, incident response procedures, and participate in on‑call rotations.

Qualifications

Strong Linux systems administration — comfortable working over SSH in production, not just dev environments.

Experience with edge or on‑prem hardware alongside cloud infrastructure.

Solid networking fundamentals: DNS, firewalls, VPNs, subnets, secure remote access.

Scripting or programming in Python, Go, or Bash for operational tooling.

Familiarity with containerization (Docker, Kubernetes a plus).

Embedded systems experience — reading firmware logs, understanding hardware‑software boundaries, and reasoning about what's happening below the OS is a meaningful edge in this role.

Deeper cloud experience (AWS infrastructure, IAM, networking, observability tooling) is a strong plus for owning the cloud side of the fleet.

Rust or C experience — we have firmware in both; being able to read and reason about low‑level code accelerates triage significantly.

#J-18808-Ljbffr

Worksite address

san francisco, CA, 94199, US

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Ready for your next step?Apply on the official website
Apply on WhatJobs ↗

Explore related searches

Current related jobs

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Space Systems Mechanical Engineer

laurel, MD

See pay details in description

Description Do you want to design and build unique space structures and spacecraft for NASA missions that enable groundbreaking scientific disco…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

System Security Engineer

laurel, MD

See pay details in description

Description Are you looking for an opportunity to utilize your technical skills to solve complex, real-world problems? If so, we're looking …

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Advanced Reentry Mission Engineer

laurel, MD

See pay details in description

Description Are you interested in hypersonic and reentry system design and prototyping? Do you want to make contributions to next generation …

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Thermal and EO/IR Modeling and Simulation Engineer

laurel, MD

See pay details in description

Description Are you looking for a unique opportunity to impact significant advances to the nation's groundbreaking integrated air and missile de…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Network Effects Engineer

laurel, MD

See pay details in description

Description Do you want to perform advanced research, development, and test & evaluation of communications systems and network technologies that…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

System Realization and Resilience Engineer

laurel, MD

See pay details in description

Description Are you passionate about applying system engineering principles to influence the development and resilience of future strategic weap…

Listing review due 2026-10-06View job