waypointjobs

CirrusLabs, LLC

Systems Reliability Engineer

jacksonville, FL

Check who can apply and the requirements below before continuing.

About this opportunity

CirrusLabs, LLC lists this Systems Reliability Engineer opportunity in jacksonville, Florida. Review the employer’s description below for duties, qualifications and application requirements.

Job description

Job Title: Systems Reliability Engineer,

Job Description:

What does a successful Systems Reliability Engineer (SRE) in Embedded Finance do on this project?

We are seeking a Systems Reliability Engineer to join the Technical Operations team supporting our Embedded Finance (EmFi) platform. In this role, you will own the reliability and resiliency of a large-scale enterprise platform, ensuring that our services remain highly available, performant and secure. You will design and implement monitoring and alerting frameworks, lead incident response and drive the root cause analysis (RCA) process to continuously improve platform stability. This is an opportunity to be a Partner in Possibility - helping our clients deliver financial services experiences that are essential to everyday life.

What you will do:

Own the reliability, resiliency and availability of the Embedded Finance platform, proactively identifying and mitigating risks to service continuity.

Design, implement and maintain comprehensive monitoring and alerting frameworks leveraging Splunk, Dynatrace, Grafana and Datadog to provide end-to-end observability across the platform.

Define and track service level objectives (SLOs), service level indicators (SLIs) and error budgets to measure and improve platform health.

Lead and participate in incident response, serving as a technical driver during remediation calls and coordinating with impacted and impacting technical and product teams.

Own and advance the root cause analysis (RCA) process — investigating incidents, documenting the sequence of events and remediating actions, and clearly identifying underlying root causes to prevent recurrence.

Ensure timely creation and management of incident tickets (e.g., ServiceNow) and accurate incident tracking, aging and reporting.

Build automation and tooling to reduce toil, improve mean time to detection (MTTD) and mean time to resolution (MTTR), and increase operational efficiency.

Collaborate with engineering, product and risk stakeholders to embed reliability best practices into the platform lifecycle.

What you will need to have:

Hands-on experience with monitoring, observability and alerting tools, specifically Splunk, Dynatrace, Grafana and Datadog.

Proven experience operating and supporting a large-scale enterprise platform environment.

Demonstrated experience with incident response and leading or contributing to root cause analysis (RCA) processes.

Strong understanding of reliability engineering principles, including availability, resiliency, monitoring and alerting best practices.

Experience with ticketing and incident management workflows (e.g., ServiceNow).

Excellent communication skills, with the ability to drive remediation efforts and collaborate across technical, product and risk teams.

What would be great to have:

Experience in financial services, payments or embedded finance environments.

Proficiency with scripting or programming languages for automation (e.g., Python, Go, Bash).

Familiarity with cloud platforms, containerization and CI/CD pipelines.

Experience defining and managing SLOs, SLIs and error budgets

#J-18808-Ljbffr

Worksite address

jacksonville, FL, 32290, US

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Ready for your next step?Apply on the official website
Apply on WhatJobs ↗

Explore related searches

Current related jobs

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Space Systems Mechanical Engineer

laurel, MD

See pay details in description

Description Do you want to design and build unique space structures and spacecraft for NASA missions that enable groundbreaking scientific disco…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

System Security Engineer

laurel, MD

See pay details in description

Description Are you looking for an opportunity to utilize your technical skills to solve complex, real-world problems? If so, we're looking …

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Advanced Reentry Mission Engineer

laurel, MD

See pay details in description

Description Are you interested in hypersonic and reentry system design and prototyping? Do you want to make contributions to next generation …

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Thermal and EO/IR Modeling and Simulation Engineer

laurel, MD

See pay details in description

Description Are you looking for a unique opportunity to impact significant advances to the nation's groundbreaking integrated air and missile de…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Network Effects Engineer

laurel, MD

See pay details in description

Description Do you want to perform advanced research, development, and test & evaluation of communications systems and network technologies that…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

System Realization and Resilience Engineer

laurel, MD

See pay details in description

Description Are you passionate about applying system engineering principles to influence the development and resilience of future strategic weap…

Listing review due 2026-10-06View job