waypointjobs

2T Consulting

Site Reliability Engineer (SRE)

Tucker, GA

Check who can apply and the requirements below before continuing.

About this opportunity

2T Consulting lists this Site Reliability Engineer (SRE) opportunity in Tucker, Georgia. Review the employer’s description below for duties, qualifications and application requirements.

Job description

We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production systems hosted on AWS. The role focuses on production support, incident management, monitoring, observability, troubleshooting, and improving system reliability and availability.

Core Responsibilities

Provide L1/L2 production support for AWS-hosted applications and infrastructure.

Monitor and troubleshoot AWS services including EC2, VPC, ALB/NLB, RDS, Lambda, and EKS.

Triage production incidents, identify root causes, restore services within defined SLAs, and escalate application defects when required.

Participate in 24/7 on-call rotations, major incident management, and post-incident reviews.

Perform application and infrastructure health checks and proactively address performance, latency, resource utilization, and availability issues.

Build and maintain monitoring and observability dashboards using CloudWatch, Dynatrace, Quantum Metric, and ThousandEyes.

Troubleshoot issues across infrastructure, networking, application, database, and Linux/Unix environments.

Support CI/CD pipelines and AWS deployment processes.

Apply reliability patterns and continuously improve system stability and operational efficiency.

Required Skills

Strong experience supporting AWS production environments.

Hands-on experience with incident management and production support.

Strong knowledge of CloudWatch, Dynatrace, Git, observability, and reliability patterns.

Experience with monitoring dashboards and operational health checks.

Strong troubleshooting and root-cause analysis skills.

Working knowledge of CI/CD, databases, and Unix/Linux.

Excellent communication and incident coordination skills.

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Ready for your next step?Apply on the official website
Apply on WhatJobs ↗

Explore related searches

Current related jobs

Annapurna Labs (u.s.) Inc.

WhatJobs

PD Engineer, Annapurna Labs

cupertino, CA

Salary not specified

As a member of the Cloud-Scale Machine Learning Acceleration team you'll be responsible for the design and optimization of Hardware in our data c…

Last received from source 2026-10-07View job

Amazon.com Services Llc - A57

WhatJobs

Senior Automation Engineer

suffolk, VA

Salary not specified

Operations is at the heart of Amazon's business. We are known for our speed, accuracy, and exceptional service. Our buildings deliver tens of tho…

Last received from source 2026-10-07View job

Annapurna Labs (u.s.) Inc.

WhatJobs

DFT Design Engineer, Machine Learning Acceleration

austin, TX

Salary not specified

Custom SoCs (System on Chip) are at the heart of AWS Machine Learning servers. As a member of the Cloud-Scale Machine Learning Acceleration team,…

Last received from source 2026-10-07View job