waypointjobs

Skill

Site Reliability Engineer

southlake, TX

Check who can apply and the requirements below before continuing.

About this opportunity

Skill lists this Site Reliability Engineer opportunity in southlake, Texas. Review the employer’s description below for duties, qualifications and application requirements.

Job description

Overview

Placement Type:

Temporary

Salary:

$ Hourly

Start Date:

Oct 8, 2026

The hiring company is a leader in the financial services industry, dedicated to empowering individuals and institutions to achieve their financial goals. They are at the forefront of innovation, constantly evolving their platforms and services to provide unparalleled value and security to their clients. This is an opportunity to join a dynamic team that is passionate about leveraging technology to drive business success and enhance the client experience.

We are seeking a highly motivated and experienced Site Reliability Engineer to join a pivotal team, partnering with our client to elevate the reliability and operational excellence of critical production systems. In this dynamic contract role, you will be instrumental in shaping the future of our infrastructure, driving automation, and ensuring seamless operations across both on-premises and cloud platforms. Your expertise will directly impact the stability, performance, and scalability of systems that serve millions, making a tangible difference in the daily lives of our clients. If you thrive on solving complex challenges, have an automation-first mindset, and are eager to contribute to a high-impact environment, this is your chance to shine.

Key Responsibilities

Develop Python-based automation solutions to reduce manual operational effort and enhance efficiency.

Automate infrastructure management across Linux, Windows, Kubernetes, cloud platforms, and cloud-native environments.

Integrate various tools and platforms through APIs and client libraries to streamline workflows.

Assist in implementing robust infrastructure automation using industry-standard technologies.

Support CI/CD automation and deployment reliability initiatives to ensure smooth and consistent releases.

Monitor and maintain production systems to consistently meet and exceed reliability and availability objectives.

Actively participate in incident response, thorough troubleshooting, and root cause analysis activities to prevent recurrence.

Develop proactive automation and operational improvements to eliminate recurring issues and improve system resilience.

Support disaster recovery, failover testing, and other operational readiness activities to ensure business continuity.

Perform in-depth performance analysis and regular system health reviews to optimize system behavior.

Build and maintain comprehensive dashboards, alerts, and monitoring solutions using Splunk, Grafana, Prometheus, or similar tools.

Improve visibility into application and infrastructure health through enhanced metrics, logs, and traces.

Investigate alerts thoroughly and identify opportunities to reduce noise and improve detection accuracy.

Explore AI/ML-driven operational improvements such as anomaly detection, intelligent alerting, and log analytics.

Assist in developing automation solutions that leverage AI to significantly improve operational efficiency.

Participate in evaluating emerging AIOps capabilities and cutting-edge observability technologies.

Required Qualifications

Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience.

3 to 5 years of hands-on experience in Site Reliability Engineering, Production Engineering, DevOps, Systems Engineering, or Platform Engineering, with a strong focus on production operations.

Must have supported production systems at scale, demonstrating a deep understanding of operational challenges in large environments.

Operations ownership, incident response, and reliability engineering experience are essential.

Strong programming skills in Python for automation and tooling development; Python must be a primary skill, not just basic scripting.

Demonstrated experience building automation tools, scripts, frameworks, or operational solutions, with the ability to provide examples of personally developed automation.

Experience supporting Kubernetes and cloud platforms (GCP, AWS, or Azure).

Familiarity with infrastructure automation and configuration management tools.

Experience with monitoring and observability platforms such as Splunk, Grafana, Prometheus, Datadog, or similar.

Understanding of Linux systems, networking, and distributed applications.

Strong analytical, troubleshooting, and problem-solving skills, with experience in critical production incident management and Root Cause Analysis (RCA).

Ability to work effectively in fast-paced, mission-critical environments, reducing operational toil through automation.

Preferred Qualifications

Experience with Terraform, Ansible, or other Infrastructure as Code solutions.

Exposure to OpenTelemetry and modern observability practices.

Experience with CI/CD pipelines and deployment automation.

Knowledge of AI/ML, AIOps, or intelligent operational tooling.

Experience supporting highly available production systems in regulated or enterprise environments.

Our eligible talent get access to amazing benefits like subsidized health, vision, and dental plans, paid sick leave, and retirement plans with a match.

Aquent is an equal-opportunity employer. We evaluate qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, veteran status, and other legally protected characteristics. We're about creating an inclusive environment-one where different backgrounds, experiences, and perspectives are valued, and everyone can contribute, grow their careers, and thrive.

#LI-LP1

#J-18808-Ljbffr

Worksite address

southlake, TX, 76092, US

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Ready for your next step?Apply on the official website
Apply on WhatJobs ↗

Explore related searches

Current related jobs

Choctaw Global

WhatJobs

Engineer Technician (McAlester, OK)

mcalester, OK

Salary not specified

Engineering Technician McAlester, OK | Full-Time | Choctaw Defense Manufacturing Build the Foundation Behind Mission-Critical Manufacturing Are y…

Last received from source 2026-10-06View job

Canon U.S.A., Inc.

WhatJobs

Field Service Engineer II - Semiconductor

hillsboro, OR

$29.20 - $43.73 hourly

Field Service Engineer II - Semiconductor US-OR-Hillsboro Job ID: 34838 Type: Full-Time # of Openings: 1 Category: Field Service CUSA OR - Rock C…

Last received from source 2026-10-06View job

Canon U.S.A., Inc.

WhatJobs

Field Service Engineer II - PVD Semiconductor

boise, ID

See pay details in description

Field Service Engineer II - PVD Semiconductor US-ID-Boise Job ID: 34587 Type: Full-Time # of Openings: 1 Category: Field Service Additional Locat…

Last received from source 2026-10-06View job

Canon U.S.A., Inc.

WhatJobs

Field Service Engineer I - Semiconductor

hillsboro, OR

See pay details in description

Field Service Engineer I - Semiconductor US-OR-Hillsboro Job ID: 34839 Type: Full-Time # of Openings: 1 Category: Field Service CUSA OR - Rock Cr…

Last received from source 2026-10-06View job

Canon U.S.A., Inc.

WhatJobs

Field Service Engineer II - PVD Semiconductor

boise, ID

See pay details in description

Field Service Engineer II - PVD Semiconductor US-ID-Boise Job ID: 34295 Type: Full-Time # of Openings: 1 Category: Field Service Additional Locat…

Last received from source 2026-10-06View job

Canon U.S.A., Inc.

WhatJobs

Field Service Engineer I - Semiconductor

san jose, CA

$27.88 - $41.75 hourly

Field Service Engineer I - Semiconductor US-CA-San Jose Job ID: 34913 Type: Full-Time # of Openings: 1 Category: Field Service CUSA San Jose Bran…

Last received from source 2026-10-06View job