Site Reliability Engineer (SRE) – Telecom Subscription PlatformsLocation: Overland Park, Kansas / Atlanta, Georgia (Local Candidates Only)
Pay Rate: $45–$50/hr
Contract: Long-Term
Schedule: 100% Onsite | 5 Days/Week
Experience: 11+ Years Required
About the Role
We are seeking an experienced Site Reliability Engineer (SRE) to support cloud-native subscription platforms within a telecom environment. The ideal candidate will ensure platform reliability, monitor production systems, automate deployments, and collaborate with engineering teams to maintain high availability and operational excellence across enterprise applications.
Responsibilities
• Monitor application health, system performance, and production environments while responding to incidents and performing Root Cause Analysis (RCA).
• Build and maintain CI/CD pipelines supporting cloud-native applications and microservices.
• Support containerized workloads using Kubernetes, Docker, and cloud platforms.
• Develop dashboards, alerts, logging standards, and operational runbooks using observability tools.
• Collaborate with engineering, DevOps, and platform teams to improve reliability, automation, and deployment efficiency.
Requirements
• 11+ years of overall IT experience with strong expertise in Site Reliability Engineering, DevOps, or Application Support.
• Experience supporting cloud-native applications, distributed systems, and enterprise production environments.
• Hands-on experience with AWS, Azure, or GCP, along with Kubernetes, Docker, Terraform, Ansible, or similar automation tools.
• Strong knowledge of CI/CD tools, including Jenkins or GitHub Actions, and scripting with Python, Shell, or Go.
• Experience with Splunk, Dynatrace, Grafana, Prometheus, OpenTelemetry, Linux administration, APIs, and microservices troubleshooting.
Position Highlights
• Role: Site Reliability Engineer (SRE)
• Employment Type: Contract
• Work Model: 100% Onsite
• Industry: Telecommunications
• Cloud Platforms: AWS, Azure, GCP
• Core Technologies: Kubernetes, Docker, Terraform, Jenkins, GitHub Actions, Splunk, Dynatrace, Grafana, Prometheus
• Methodology: SRE, DevOps, Agile
• Support Model: 24/7 Production Support Rotation
Join a collaborative engineering team where you'll strengthen the reliability, scalability, and performance of enterprise telecom platforms while driving automation and operational excellence across cloud-native environments.