1

Sre Product Manager Jobs (NOW HIRING)

Site Reliability Engineer - SRE

Atlanta, GA · On-site +1

$54.25 - $72/hr

As a Staff Software Engineer, you will be a core player on the product team and are expected to build and grow the skillsets of the more junior Engineers. As a Staff Site Reliability Engineer you ...

Site Reliability Engineer - SRE

Atlanta, GA

$54.25 - $72/hr

As a Staff Software Engineer, you will be a core player on the product team and are expected to build and grow the skillsets of the more junior Engineers. As a Staff Site Reliability Engineer you ...

Site Reliability Engineer

Brentwood, TN · On-site +1

$157K - $190K/yr

The Site Reliability Engineer is accountable for delivering high-quality software while developing people, improving engineering practices, and partnering closely with Product Management and business ...

$42 - $55.75/hr

Knowledge or experience with agentic flows/development and with AI-assisted production/reliability ... On the SRE team, you'll have the opportunity to manage the complex challenges of scale which are ...

New

Site Reliability Engineer (SRE)

Red Oak, GA · On-site

$54 - $71.75/hr

We are seeking an experienced Site Reliability Engineer (SRE) to support and maintain production ... The role focuses on production support, incident management, monitoring, observability ...

Site Reliability Engineer - SRE

Atlanta, GA · On-site

$54.75 - $72.75/hr

As a Staff Software Engineer, you will be a core player on the product team and are expected to build and grow the skillsets of the more junior Engineers. As a Staff Site Reliability Engineer you ...

Site Reliability Engineer (SRE)

Miami, FL · On-site

$54.50 - $72.50/hr

* SRE & Automation: 6-8 years of experience as a Site Reliability Engineer (SRE) or DevOps Engineer with a heavy focus on building automation tools. * AI/LLM Integration: Must have experience (or ...

$42 - $55.75/hr

The Senior Microsoft SRE will provide site reliability, production operations, and cloud platform engineering support for Microsoft Azure capabilities within the CRM, Integration, and Platform ...

Site Reliability Engineer

Saint George, UT · On-site

$53.75 - $71.50/hr

The Site Reliability Engineer works as part of a team to analyze, troubleshoot, deploy, monitor ... increase product reliability and organizational efficiency * Manages solutions and ensures ...

$42 - $55.75/hr

... management tools. * Help to define the DR plan needed for our critical apps. * Help to develop a chaos testing process. * Participate in an on-call rotation and support production Incidents. SRE ...

Site Reliability Engineer III Location: Pennington, NJ Duration: Contract - 7 months Pay Range: $73 ... Incident management, production support, and root cause analysis expertise. * Experience ...

New

$42 - $55.75/hr

Manage incident response and post-mortems * Improve system reliability and performance * Automate operational processes * Define and track SLOs/SLIs Requirements * 4+ years of SRE or similar ...

Site Reliability Engineer (SRE)

Seattle, WA · On-site

$65 - $86.25/hr

Manage incident response and post-mortems * Improve system reliability and performance * Automate operational processes * Define and track SLOs/SLIs Requirements * 4+ years of SRE or similar ...

Showing results 41-60

Sre Product Manager information

See salary details

$51.5K

$159.4K

$197K

How much do sre product manager jobs pay per year?

As of Sep 13, 2026, the average yearly pay for sre product manager in the United States is $159,405.00, according to ZipRecruiter salary data. Most workers in this role earn between $141,000.00 and $197,000.00 per year, depending on experience, location, and employer.

What is an SRE Product Manager?

An SRE (Site Reliability Engineering) Product Manager is a professional who bridges the gap between site reliability engineering teams and product development. They are responsible for defining and prioritizing reliability-focused product features, working to ensure systems are scalable, reliable, and meet user needs. SRE Product Managers collaborate with engineers to advocate for reliability in the product roadmap, track service-level objectives, and drive incident management improvements. Their role is crucial in balancing innovation speed with system stability and user satisfaction.

How does an SRE Product Manager typically balance reliability goals with rapid product development?

An SRE Product Manager collaborates closely with both Site Reliability Engineering and development teams to ensure that new features are delivered without compromising system stability. This involves setting clear service level objectives (SLOs), prioritizing work that addresses technical debt, and facilitating communication between stakeholders. The role often requires making trade-offs and advocating for investments in reliability as part of product planning, while also supporting the team's agility and innovation. Effective SRE Product Managers use data-driven insights and incident postmortems to guide decision-making and align reliability initiatives with business goals.

What are the key skills and qualifications needed to thrive as an SRE Product Manager, and why are they important?

To thrive as an SRE Product Manager, you need a solid understanding of site reliability engineering principles, product management methodologies, and experience with cloud infrastructure or DevOps, often supported by a degree in computer science or a related field. Familiarity with tools such as Kubernetes, Prometheus, Jira, and CI/CD systems, as well as certifications like PMP or AWS Certified Solutions Architect, are commonly important. Outstanding communication, stakeholder management, and problem-solving skills help bridge the gap between engineering teams and business objectives. These abilities ensure the delivery of reliable, scalable products that align with both technical and customer requirements.

What are popular job titles related to Sre Product Manager jobs?

For Sre Product Manager jobs, the most frequently searched job titles are:

Infographic showing various Sre Product Manager job openings in the United States as of August 2026, with employment types broken down into 88% Full Time, 11% Part Time, and 1% Contract. Highlights an 80% Physical, 2% Hybrid, and 18% Remote job distribution, with an average salary of $159,405 per year, or $76.6 per hour.

Site Reliability Engineer

Atlanta, GA • On-site

Intercontinental Exchange Holdings, Inc.
Finance and Insurance • 5 - 10K employees

$54.75 - $72.75/hr

Full-time

Posted 5 days ago


Job description

Overview
Job Purpose
At Intercontinental Exchange (NYSE:ICE), we engineer technology, exchanges and clearing houses that connect companies around the world to global capital and derivative markets. With a leading-edge approach to developing technology platforms, we have built market infrastructure in all major trading centers, offering customers the ability to manage risk and make informed decisions globally. By leveraging our core strengths in technology, we continue to identify new ways to serve our customers and transform global markets. We're looking for motivated, results-oriented people to join our team.
We are seeking a Site Reliability Engineer II to bring 3+ years of hands-on experience to our SRE team, operating with significant autonomy to improve platform reliability, drive automation, and mentor junior engineers. The ideal candidate contributes meaningfully to platform release cycles, leads smaller projects, and actively shapes the team's approach to observability, incident response, and service design in ICE's 24x7 production environment.
Responsibilities
  • Employ advanced troubleshooting and root-cause analysis to improve availability, performance, and security of IMT and platform services
  • Collaborate with Product and Engineering teams to plan and deploy product releases with operational rigor and quality gates
  • Work with Engineering leadership to build and evolve shared services meeting the requirements of platform and application teams
  • Design and implement proactive monitoring, alerting, trend analysis, and self-healing automation
  • Resolve product and service defects, infrastructure issues, and operational changes with increasing independence
  • Implement automated tests, automated deployments, and operational tooling across the SRE toolchain
  • Ensure services are designed with 24x7 availability and operational readiness and rigor
  • Lead smaller projects and provide status updates to management and stakeholders
  • Mentor SRE I engineers and contribute actively to team training and knowledge-sharing
  • Partner with application and platform teams to identify critical workflows and build automated health checks that run post-deployment and during incidents to accelerate root-cause identification
  • Design and build AI-assisted automated diagnosis jobs that correlate signals across monitoring and alerting platforms to reduce Mean Time to Resolution (MTTR) for production incidents
  • Build and maintain automation pipelines (e.g., Rundeck, Jenkins) that integrate with AI/LLM tooling to drive efficiency gains in observability, runbook execution, and incident triage
  • Develop and tune AWS CloudWatch metrics, alarms, and dashboards, instrument services using OpenTelemetry/Alloy, and build observability visualizations in Grafana; integrate alerting and event correlation workflows across PagerDuty, BigPanda, and Splunk to ensure timely, actionable incident notification

Knowledge and Experience
  • Bachelor's degree in Computer Science, Engineering, or equivalent experience
  • 3+ years of experience in a site reliability, production engineering, or software operations role
  • Proven technical skills with strong personal initiative and consistent delivery of important work
  • Excellent teamwork with active involvement in training and mentoring
  • Ability to prioritize and execute without direct management guidance
  • Strong understanding of ICE Core Competencies

Preferred Knowledge and Experience
  • Experience in financial services technology, mortgage platforms, or exchange infrastructure
  • Familiarity with SRE principles including SLI, SLO, and error budget management
  • Exposure to Terraform, Chef, Ansible, or equivalent infrastructure automation frameworks
  • Hands-on experience with AWS observability services, CloudWatch, Grafana, OpenTelemetry/Alloy, Splunk, BigPanda, PagerDuty, and job orchestration/automation platforms such as Rundeck and Jenkins
  • Practical experience integrating AI/LLM-based tooling into operational workflows to automate diagnosis, reduce manual triage, and improve incident response efficiency

#LI-JM1
-
Intercontinental Exchange, Inc. is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to legally protected characteristics.