1

Observability Site Reliability Engineer Jobs in Minnesota

Provide SRE production support for Mainframe applications, APIs, Web Services, and Microservices ... Build and enhance monitoring, alerting, dashboards, observability, and self-healing capabilities.

Site Reliability Engineer

Minnetonka, MN

$58 - $77.25/hr

Provide SRE production support for Mainframe applications, APIs, Web Services, and Microservices ... Build and enhance monitoring, alerting, dashboards, observability, and self-healing capabilities.

Lead Platform Reliability Engineer

Minneapolis, MN · On-site

$107K - $134K/yr

As part of our Platform Reliability Engineering (PRE) team, you will apply modern Site Reliability Engineering (SRE) practices to improve the availability, resiliency, observability, scalability, and ...

Site Reliability Engineer

Minneapolis, MN · On-site

$130K - $150K/yr

The Site Reliability Engineer will be asked to solve various technical challenges as well as participate in many strategic discussions for development and deployment. Essential Job Duties * Improve ...

Site Reliability Engineer

Minneapolis, MN · On-site

$130K - $150K/yr

The Site Reliability Engineer will be asked to solve various technical challenges as well as participate in many strategic discussions for development and deployment. Essential Job Duties * Improve ...

Site Reliability Engineer.

Saint Paul, MN · On-site

$57.75 - $76.50/hr

This is the CI/CD and DevOps experience in the feedback. 5. Migration / remediation experience 6. MSSQL 7. VMware vSphere & associated tools 8. Windows Application Monitoring tool Optional · .net ...

next page

Showing results 1-20

Observability Site Reliability Engineer information

What engineer makes $500,000 a year?

A senior or principal Site Reliability Engineer (SRE) or Observability Engineer with extensive experience, specialized skills, and working at large tech companies can earn $500,000 or more annually. Compensation often includes base salary, bonuses, and stock options, especially in high-demand markets and organizations with complex infrastructure.

Is AI replacing SRE?

AI is augmenting the work of Site Reliability Engineers (SREs) by automating tasks such as monitoring, incident detection, and response. However, SREs are still essential for designing systems, managing complex issues, and making strategic decisions that require human judgment. AI tools are considered complementary rather than replacements for SREs' expertise and problem-solving skills.

What engineers make $200,000 a year?

Senior Site Reliability Engineers and Observability Engineers with extensive experience, advanced skills in cloud platforms, automation, and monitoring tools can earn $200,000 or more annually. High compensation often correlates with working at large tech companies, possessing specialized certifications, and managing complex, scalable systems.

What is the difference between Observability Site Reliability Engineer vs Monitoring Engineer?

AspectObservability Site Reliability EngineerMonitoring Engineer
FocusEnsuring system reliability through observability, automation, and incident responseImplementing and managing monitoring tools and dashboards
SkillsCloud platforms, scripting, incident management, observability toolsMonitoring tools, alerting systems, data analysis
Work EnvironmentDevOps teams, cloud infrastructure, large-scale systemsOperations teams, infrastructure monitoring

While both roles involve system health, the Observability Site Reliability Engineer focuses on comprehensive system reliability using observability practices, whereas Monitoring Engineers primarily manage monitoring tools and alerts. The SRE role emphasizes automation, incident response, and system resilience, making it broader in scope.

What engineers make $300,000 a year?

Senior Site Reliability Engineers and Observability Engineers with extensive experience, advanced skills in cloud platforms, automation, and monitoring tools can earn $300,000 or more annually. High compensation often correlates with working at large tech companies, possessing specialized certifications, and taking on leadership or highly technical roles.
What job categories do people searching Observability Site Reliability Engineer jobs in Minnesota look for? The top searched job categories for Observability Site Reliability Engineer jobs in Minnesota are:
What cities in Minnesota are hiring for Observability Site Reliability Engineer jobs? Cities in Minnesota with the most Observability Site Reliability Engineer job openings:
Senior Site Reliability Engineer

Senior Site Reliability Engineer

Royal Bank of Canada

Minneapolis, MN • On-site

$59.50 - $79/hr

Full-time

Posted 13 days ago


Job description

Job Description

RBC Wealth Management Technology is seeking a Senior Site Reliability Engineer to join its Wealth Management SRE Team. This team is responsible for ensuring the performance, availability, resilience, and operational excellence of critical applications and platforms that support the wealth management business. Working at the intersection of software engineering, cloud-native operations, observability, and automation, the team plays a central role in delivering reliable digital services for both internal users and clients.

As a Senior Site Reliability Engineer, you will bring an engineering-first mindset, strong operational judgment, and a passion for automation to improve system reliability at scale. You will work closely with development, infrastructure, platform, and support teams to build modern observability practices, improve incident response, strengthen reliability engineering standards, and drive the evolution toward intelligent, self-healing operations.

This role is ideal for a hands-on engineer who is equally comfortable improving production resilience, building automation, defining service-level objectives, and shaping the future of AI-enhanced operations. You will help design and implement scalable SRE solutions across the technology estate using tools and platforms such as Elasticsearch, Ansible, GitHub Actions, Dynatrace, PagerDuty, Moogsoft, Kubernetes, OpenShift, Kafka, and emerging AIOps capabilities.

What will you do?

  • Build and enhance the SRE product base - Develop intelligent monitoring, alerting, reliability testing, anomaly detection, and automated remediation capabilities.

  • Implement modern observability practices - Deploy metrics, logs, traces, dashboards, and actionable alerting across supported applications.

  • Design ML-based anomaly detection and self-healing solutions - Shift from reactive to predictive operations with automated issue remediation and appropriate governance controls.

  • Standardize telemetry and instrumentation - Improve visibility, coverage, and correlation of operational signals across platforms.

  • Automate operational workflows - Use Ansible, GitHub Actions, and scripting (Bash, Python, PowerShell) to streamline platform tasks and develop custom tooling.

  • Define and track service health metrics - Establish and improve SLIs, SLOs, error budgets, and evolve runbooks into automation-first remediation patterns.

  • Partner with development teams - Ensure applications meet reliability and performance standards before and after deployment through close collaboration.

  • Lead incident and problem management - Troubleshoot production issues across all layers, participate in on-call rotation, and drive root cause analysis and corrective actions.

  • Drive continuous improvement - Identify opportunities to simplify, automate, and modernize operations using engineering and AI-driven approaches

What do you need to succeed?

Must-have

  • 5+ years of experience in Site Reliability Engineering, Production Engineering, DevOps, Platform Engineering, or Systems Engineering roles with strong operational depth.

  • Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience.

  • Strong experience with infrastructure automation and configuration management, particularly Ansible.

  • Strong scripting and automation skills in Bash, Python, PowerShell, or similar languages.

  • Hands-on experience with modern reliability and observability tooling such as Elasticsearch, Dynatrace, GitHub, Kubernetes, OpenShift, Kafka, PagerDuty, Moogsoft, or related platforms.

  • Strong understanding of production operations, incident management, root cause analysis, and reliability engineering practices.

  • Experience defining and operating SLIs, SLOs, alerting strategies, and service health metrics.

  • Knowledge of cloud-native and distributed systems concepts, including resiliency, scalability, fault isolation, and performance tuning.

  • Understanding of AIOps, AI/ML concepts, or intelligent automation as applied to observability and operations.

  • Ability to work across teams, influence engineering practices, and communicate clearly with technical and non-technical stakeholders.

Nice-to-have

  • Experience in financial services, wealth management, banking, insurance, or other highly regulated environments.

  • Experience with OpenTelemetry and telemetry standardization across distributed systems.

  • Hands-on experience with Prometheus, Grafana, Splunk, Catchpoint, Azure Automation, or similar SRE and observability platforms.

  • Experience with CI/CD and developer platform tools such as Jenkins, Artifactory, and Vault.

  • Familiarity with containerization and cloud platform patterns, including Docker and Kubernetes-based deployments.

  • Experience building or operating anomaly detection, predictive alerting, or self-healing automation solutions.

  • Familiarity with AI governance, model validation, and operational controls in regulated environments.

What's in it for you?

We thrive on the challenge to be our best, progressive thinking to keep growing, and working together to deliver trusted advice to help our clients thrive and communities prosper. We care about each other, reaching our potential, making a difference to our communities, and achieving success that is mutual.

  • A comprehensive Total Rewards Program including bonuses and flexible benefits, competitive compensation, commissions, and stock where applicable

  • Leaders who support your development through coaching and managing opportunities

  • Ability to make a difference and lasting impact

  • Work in a dynamic, collaborative, progressive, and high-performing team

  • A world-class training program in financial services

The expected salary range for this particular position is $80,000-$140,000, depending on your experience, skills, and registration status, market conditions and business needs.

You have the potential to earn more through RBC's discretionary variable compensation program which gives you an opportunity to increase your total compensation, provided the business meets its performance targets and you meet your individual goals.

RBC's compensation philosophy and principles recognize the importance of a highly qualified global workforce and plays a critical role in attracting, engaging and retaining talent that:

  • Drives RBC's high-performance culture.

  • Enables collective achievement of our strategic goals.

  • Generates sustainable shareholder returns and above market shareholder value.

#LI-POST

Job Skills

Agile Methodology, Application Infrastructure, Group Problem Solving, IT Automation, IT Monitoring, Operations Support, Production Support, Software Development Life Cycle (SDLC), Software Engineering, Software Product Technical Knowledge, System Applications, Systems Software

Additional Job Details

Address:

250 NICOLLET MALL:MINNEAPOLIS

City:

Minneapolis

Country:

United States of America

Work hours/week:

40

Employment Type:

Full time

Platform:

TECHNOLOGY AND OPERATIONS

Job Type:

Regular

Pay Type:

Salaried

Posted Date:

2026-07-13

Application Deadline:

2026-08-07

Note: Applications will be accepted until 11:59 PM on the day prior to the application deadline date above

Our Employment Opportunities

At RBC, we are guided by living shared values of Client First, Integrity, Collaboration, Respect and Excellence and winning together as One RBC. We believe an inclusive workplace that has diverse perspectives is core to our continued growth as one of the largest and most successful banks in the world. Maintaining a workplace where our employees feel supported to perform at their best, effectively collaborate, drive innovation, and grow professionally helps to bring our Purpose to life and create value for our clients and communities. RBC strives to deliver this through policies and programs intended to foster a workplace based on respect, belonging and opportunity for all.

Join our Talent Community
Stay in-the-know about great career opportunities at RBC. Sign up and get customized info on our latest jobs, career tips and Recruitment events that matter to you.
Expand your limits and create a new future together at RBC. Find out how we use our passion and drive to enhance the well-being of our clients and communities at jobs.rbc.com.

RBC is presently inviting candidates to apply for this existing vacancy. Applying to this posting allows you to express your interest in this current career opportunity at RBC. Qualified applicants may be contacted to review their resume in more detail.

Employment Type: FULL_TIME