waypointjobs

Compunnel, Inc.

Reliability Engineer

town of texas, WI

Check who can apply and the requirements below before continuing.

About this opportunity

Compunnel, Inc. lists this Reliability Engineer opportunity in town of texas, Wisconsin. Review the employer’s description below for duties, qualifications and application requirements.

Job description

JOB SUMMARY

The Reliability Engineering group within Enterprise Infrastructure combines Operations Excellence with the Development Experience to deliver services at high scale, high availability with resilience by using automation and Infrastructure Code. We build reliability into our ecosystem by applying standards in Resiliency Engineering, Automation, Observability & Chaos Testing. Additionally, this role contributes to enterprise backup and recovery capabilities including automation of recovery workflows, rehoused recovery into alternate datacenters, and testing of recovery processes. We are looking for a systems thinking, reliability engineer who has helped teams scale through production insight, data and backup recovery, operational automation, developer guidance, real-time metrics, automation, automation, automation. Crafting scalable solutions and automation to monitor the health and establish signals to drive understanding of our Container Platform environments. Strengthening operational processes with client’s support and incident management teams for our cloud ecosystem Working with client’s and cloud service provider product teams and driving ongoing reliability improvements in their Kubernetes service offerings. Anticipating, discovering through ongoing interaction with, and prioritizing client / partner needs to serve as their voice and guide execution of the team.

Key Responsibilities

Automate recovery workflows

Rehouse recovery into alternate datacenters

Test recovery processes

Advance enterprise resiliency through improved recovery capabilities

Reduce recovery time via automation

Enable rehoused recovery into new datacenters

Strengthen platform reliability through data protection design

Craft scalable solutions and automation to monitor the health and establish signals to drive understanding of Container Platform environments

Strengthen operational processes with client’s support and incident management teams for cloud ecosystem

Drive ongoing reliability improvements in their Kubernetes service offerings

Prioritize client / partner needs to serve as their voice and guide execution of the team

Required Qualifications

Bachelor’s Degree or equivalent experience in a technology related field (e.g. Computer Science, Engineering, etc.) required.

Production experience running Cloud and on‑prem Storage workloads at scale

Experience managing and maintaining Kubernetes Clusters on EKS/AKS and RKS.

Demonstrates a drive for continuous improvement and enjoys tackling complex problems.

Experience managing and interpreting large datasets using query languages and visualization tools (PowerBI/tableau)

Experience in software development with Python, NodeJS, or Java with a focus on SDLC and automation

5 -7 years of hands‑on experience deploying and/or supporting highly distributed multi‑tiered systems at scale.

Experience building and deploying Docker images including Docker Compose

Hands‑on experience with Jenkins Core, including authoring and maintaining declarative CI/CD pipelines and libraries

Experience with distributed version control systems, Git preferred

Experience crafting and maintaining logging, monitoring, and alerting capabilities using tools like Datadog and Splunk

Practical experience in building cloud hosted and native applications for the enterprise.

Maintains a deep understanding of a wide variety of AWS/Azure services that support reliability, observability, and automation/orchestration.

Experience in incident/crisis management and supporting critically important applications

Hard on experience with one or more observability tools (Prometheus, Grafana, ELK/OpenSearch, OpenTelemetry, Datadog, etc.)

Ability to automate with various scripting languages (Python, Shell scripting, etc.)

Experience managing systems using infrastructure as code tools (IAM, ARM, Terraform, Chef)

Preferred Qualifications

Go, Angular, Python, JavaScript, AWS, RESTful services, Ruby, MVC, Jenkins CI/CD, Configuration Automation (Chef, Ansible)

Bootstrap, HTML/CSS, Shell Scripting, messaging frameworks (MQ), Service Oriented/Micro-service Architectures, OpenStack, Relational Databases (PostgreSQL)

Comfortable working in both Public and private cloud environments

#J-18808-Ljbffr

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Ready for your next step?Apply on the official website
Apply on WhatJobs ↗

Explore related searches

Current related jobs

L3Harris Technologies

WhatJobs

Lead, Electrical Engineer (Power)

palm bay, FL

Salary not specified

L3Harris is dedicated to recruiting and developing high-performing talent who are passionate about what they do. Our employees are unified in a s…

Last received from source 2026-10-07View job

L3Harris Technologies

WhatJobs

Lead, Project Engineer

melbourne, FL

Salary not specified

L3Harris is dedicated to recruiting and developing high-performing talent who are passionate about what they do. Our employees are unified in a s…

Last received from source 2026-10-07View job

L3Harris Technologies

WhatJobs

Scientist, Systems Engineer

palm bay, FL

Salary not specified

L3Harris is dedicated to recruiting and developing high-performing talent who are passionate about what they do. Our employees are unified in a s…

Last received from source 2026-10-07View job

L3Harris Technologies

WhatJobs

Senior Specialist, Systems Engineer

palm bay, FL

Salary not specified

L3Harris is dedicated to recruiting and developing high-performing talent who are passionate about what they do. Our employees are unified in a s…

Last received from source 2026-10-07View job