1

Sre Engineer Devops Engineer Jobs in Chicago, IL

Sr. Site Reliability Engineer (SRE)

Chicago, IL · On-site

$58.75 - $78/hr

... DevOps, or infrastructure engineering roles with proven experience operating production ... SRE principles including SLIs/SLOs/SLAs, error budgets, incident management, and blameless ...

Senior DevOps/SRE Engineer

Chicago, IL · On-site

$140K - $170K/yr

BA/BS, in a related technical field; or the equivalent in education and work experience * 8+ years of experience in DevOps, SRE, platform engineering, or similar roles supporting application teams ...

Senior DevOps/SRE Engineer

Chicago, IL · On-site

$140K - $170K/yr

BA/BS, in a related technical field; or the equivalent in education and work experience * 8+ years of experience in DevOps, SRE, platform engineering, or similar roles supporting application teams ...

Senior DevOps/SRE Engineer

Chicago, IL · On-site

$140K - $170K/yr

BA/BS, in a related technical field; or the equivalent in education and work experience * 8+ years of experience in DevOps, SRE, platform engineering, or similar roles supporting application teams ...

Senior Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

Chicago, IL (Hybrid 3X a week) Senior Site Reliability Engineer (SRE) Chicago, IL (Hybrid ... DevOps, or cloud infrastructure. * Bachelor's degree in Computer Science, Computer Engineering ...

Senior Site Reliability Engineer

Chicago, IL · On-site

$58.75 - $78/hr

Chicago, IL (Hybrid 3X a week) Senior Site Reliability Engineer (SRE) Chicago, IL (Hybrid ... DevOps, or cloud infrastructure. * Bachelor's degree in Computer Science, Computer Engineering ...

DevOps Engineer

Chicago, IL · On-site

$54.50 - $74.50/hr

Required Skills & Experience * 3-6 years of DevOps or SRE experience. * Strong hands‑on with AWS, Azure, or GCP. * Expertise in CI/CD tools. * Infrastructure as Code (Terraform, ARM, CloudFormation)

DevOps Engineer

Chicago, IL

$54.50 - $74.50/hr

Required Skills & Experience: * 3 -6 years of DevOps or SRE experience. * S trong hands-on with AWS, Azure, or GCP. * E xpertise in CI/CD tools. * I nfrastructure as Code (Terraform, ARM ...

DevOps Engineer

Chicago, IL · On-site

$54.50 - $74.50/hr

Required Skills & Experience: - 3-6 years of DevOps or SRE experience. - Strong hands-on with AWS, Azure, or GCP. - Expertise in CI/CD tools. - Infrastructure as Code (Terraform, ARM, CloudFormation ...

DevOps / Cloud Engineer

Chicago, IL · On-site

$73K - $174K/yr

Required Qualification * 4-8 years of experience in DevOps, Cloud Engineering, Site Reliability Engineering (SRE), or Infrastructure Engineering. * Strong hands-on experience with AWS services such ...

Sr Site Reliability Engineer

Wheeling, IL · On-site

$150K - $185K/yr

Author and troubleshoot Azure DevOps pipelines; support teams with deployment visibility, change ... Core SRE Experience By providing your phone number, you consent to: (1) receive automated text ...

New

DevOps / Cloud Engineer

Chicago, IL · On-site

$73K - $174K/yr

Required Qualification4-8 years of experience in DevOps, Cloud Engineering, Site Reliability Engineering (SRE), or Infrastructure Engineering.Strong hands-on experience with AWS services such as EC2 ...

Staff SRE

Chicago, IL · Hybrid

$58.75 - $78/hr

Staff Site Reliability Engineer (Technical Lead) - Platform Engineering Location: Chicago, IL ... DevOps, Systems Engineering, or Software Engineering. * Experience as a Technical Lead, Staff ...

DevOps / Cloud Engineer

Chicago, IL · On-site

$73K - $174K/yr

Required Qualification4-8 years of experience in DevOps, Cloud Engineering, Site Reliability Engineering (SRE), or Infrastructure Engineering.Strong hands-on experience with AWS services such as EC2 ...

Site Reliability Engineer

Chicago, IL · On-site

$100K - $120K/yr

Partners with the larger Cloud Operations, SRE, Engineering teams, and the business-at-large to advance our SaaS platforms. * Other duties as assigned. Qualifications * Bachelor's degree in Computer ...

New

Site Reliability Engineer

Riverwoods, IL · On-site

$59.25 - $78.75/hr

We are looking for a Site Reliability Engineer to join our growing team in Riverwoods, IL United ... Experience with DevOps, CI/CD tools. * Good experience of Linux, AWS Cloud and on Prem deployments.

Showing results 21-40

Sre Engineer Devops Engineer information

See Chicago, IL salary details

$11

$65

$94

How much do sre engineer devops engineer jobs pay per hour?

As of Sep 15, 2026, the average hourly pay for sre engineer devops engineer in Chicago, IL is $65.66, according to ZipRecruiter salary data. Most workers in this role earn between $56.44 and $75.05 per hour, depending on experience, location, and employer.

What is an SRE engineer or DevOps engineer?

SRE (Site Reliability Engineer) and DevOps Engineer are roles focused on ensuring the reliability, scalability, and efficiency of software systems. SREs typically use software engineering approaches to automate IT operations tasks, improve system reliability, and manage incidents. DevOps Engineers focus on bridging the gap between development and operations by automating deployments, managing CI/CD pipelines, and fostering a collaborative culture. Both roles share similar goals but differ in focus: SRE emphasizes reliability and incident management, while DevOps centers around streamlining development and operational processes.

What skills and qualifications are needed to thrive as an SRE engineer or DevOps engineer?

To thrive as an SRE Engineer or DevOps Engineer, you need strong knowledge of system administration, cloud infrastructure, automation, and scripting languages, typically backed by a degree in computer science or related experience. Familiarity with tools such as Kubernetes, Docker, Jenkins, Terraform, and monitoring systems, as well as certifications like AWS Certified DevOps Engineer or Google Professional DevOps Engineer, are highly valuable. Excellent problem-solving, collaboration, and communication skills help you work effectively across teams and respond to incidents rapidly. These combined skills and qualities ensure the reliability, scalability, and efficiency of critical production systems in dynamic technology environments.

What are common challenges SRE engineers or DevOps engineers face when balancing reliability with rapid deployment?

SRE/DevOps Engineers often encounter the challenge of maintaining system reliability and uptime while supporting rapid development and frequent deployments. Balancing these priorities requires implementing automated testing, robust monitoring, and efficient rollback strategies to minimize risk. Effective collaboration with development and operations teams is essential to ensure that new features are released safely without compromising system stability. Continuous communication and clear incident response processes also help address issues quickly and maintain service quality.

What is the difference between Sre Engineer Devops Engineer vs Cloud Engineer?

AspectSre Engineer Devops EngineerCloud Engineer
Primary FocusEnsuring system reliability, automation, and performance of servicesDesigning, deploying, and managing cloud infrastructure and services
Skills & CertificationsLinux, scripting, CI/CD, monitoring tools, often certifications like AWS, Azure, GCPCloud platforms (AWS, Azure, GCP), cloud architecture, certifications like AWS Certified Solutions Architect
Work EnvironmentOperations, development, and infrastructure teams in tech companiesCloud service providers, cloud-focused teams, DevOps teams

While Sre Engineers Devops Engineers focus on system reliability, automation, and performance, Cloud Engineers specialize in designing and managing cloud infrastructure. Both roles often collaborate but have distinct core responsibilities within tech organizations.

Are SRE and DevOps engineers the same?

SRE (Site Reliability Engineer) and DevOps engineers both focus on improving software deployment and system reliability, but their roles differ; SREs typically emphasize reliability, monitoring, and automation using tools like Prometheus and Kubernetes, while DevOps engineers focus on integrating development and operations processes to enable continuous delivery. Both roles require strong scripting skills and collaboration across teams, but SREs often have a more specialized focus on system stability and incident response.

What cities near Chicago, IL are hiring for Sre Engineer Devops Engineer jobs?

Cities near Chicago, IL with the most Sre Engineer Devops Engineer job openings:

Sr. Site Reliability Engineer (SRE)

Chicago, IL • On-site

$58.75 - $78/hr

Full-time

Re-posted 3 days ago


Job description

Job Summary:
Moonlite AI delivers high-performance AI infrastructure for organizations running intensive computational research and large-scale model training. The Sr. Site Reliability Engineer will be responsible for building and operating production-grade AI infrastructure, focusing on Kubernetes expertise to ensure reliability and performance across multiple regions.
Responsibilities:
• Design, build, and operate production Kubernetes clusters on bare-metal infrastructure – including cluster bootstrapping, control plane architecture, etcd management, and scaling strategies for high-performance compute workloads.
• Implement and operate custom Kubernetes networking solutions with SR-IOV for high-performance GPU interconnects, multi-tenancy isolation and advanced networking policies. Configure CNI plugins and network segmentation for research workloads.
• Develop and maintain custom Kubernetes operators and controllers for bare-metal provisioning, infrastructure lifecycle management, and resource orchestration across compute, storage, and networking domains.
• Deploy and optimize NVIDIA GPU operators, device plugins, and other custom scheduling logic for GPU workload placement and utilization optimization.
• Build deep integrations between Kubernetes and underlying infrastructure including CSI drivers for storage, custom admission controllers for policy enforcement, and scheduling extensions for specialized hardware placement.
• Design and implement automation using Terraform, Ansible, Helm, and custom operators to orchestrate infrastructure workflows and enable deployments across multiple regions.
• Manage production bare-metal infrastructure across multiple regions. Build systems ensuring high availability, fault tolerance, and graceful degradation – establishing SLIs, SLOs, and monitoring to meet enterprise reliability commitments.
• Build comprehensive monitoring, logging, and alerting using Prometheus, Grafana, and ELK stack. Lead incident response, conduct postmortems, and implement preventative measures to improve reliability and reduce MTTR.
• Identify and resolve performance bottlenecks across infrastructure domains. Monitor utilization trends, forecast capacity needs, and optimize resource allocation for various workloads.
Qualifications:
Required:
• 5+ years in SRE, DevOps, or infrastructure engineering roles with proven experience operating production infrastructure at scale.
• Deep hands-on experience building and operating production Kubernetes clusters on bare-metal infrastructure – not just deploying workloads in managed clusters. Must understand cluster bootstrapping, control plane architecture, etcd operations, and scaling strategies.
• Strong understanding of Kubernetes internals including custom resource definitions (CRDs), operators, controllers, admission webhooks, and scheduling. Experience integrating storage (CSI drivers), networking (CNI, SR-IOV), and specialized hardware (GPU device plugins) with Kubernetes.
• Strong fundamentals in Linux systems administration, performance tuning, troubleshooting, and automation in production environments.
• Proficiency with infrastructure-as-code tools (Terraform, Ansible, Helm) and building automation to reduce operational overhead.
• Solid understanding of networking concepts including IPAM, DNS, DHCP, VLAN/VXLAN, routing, load balancing, and experience troubleshooting network issues in production.
• Experience building and maintaining comprehensive monitoring solutions using tools like Prometheus, Grafana, and centralized logging systems.
• Understanding of SRE principles including SLIs/SLOs/SLAs, error budgets, incident management, and blameless postmortems.
• Strong scripting skills in Go, Python, or Bash for automation, tooling development, and operational efficiency.
• Demonstrated ability to troubleshoot complex issues under pressure, manage incidents effectively, and communicate clearly during outages.
• Excellent communication skills and ability to work across teams including systems engineers, network engineers, and software developers.
Preferred:
• Experience building custom Kubernetes operators or controllers for infrastructure orchestration
• Deep familiarity with Kubernetes networking (Calico, Cilium, Multus), service mesh technologies, and network policy management
• Experience with GPU workload orchestration including NVIDIA GPU Operator, MIG, time-slicing, and device plugins
• Background with advanced Kubernetes features including custom schedulers, admission controllers, and API server extensions
• Experience with Kubernetes cluster federation or multi-cluster management
• Knowledge of high-performance networking technologies (InfiniBand, RDMA, RoCE) and their integration with Kubernetes
• Experience with enterprise storage systems (VAST, Lightbits, Ceph, or similar)
• Familiarity with configuration management at scale and GitOps practices
• Understanding of security best practices for Kubernetes and bare-metal infrastructure
• Experience operating infrastructure in regulated industries or co-located data center environments
• Background supporting research institutions, technical computing environments, or enterprise AI infrastructure
Company:
Moonlite AI is a technology company. Founded in 2024, the company is headquartered in Chicago, USA, with a team of 2-10 employees. The company is currently Early Stage.