1

Observability Site Reliability Engineer Jobs in Iowa

Lead Reliability Engineer

Bettendorf, IA · On-site

$91K - $115K/yr

Performance of criticality analysis and determination of risk priority numbers for the site ... Bachelor's degree in Mechanical Engineering, Electrical Engineering, Reliability Engineering or a ...

Observability & Monitoring: Design and implement monitoring frameworks to govern service-oriented ... Experience: 6+ years of experience in SRE, DevOps, or Systems Engineering roles, with a deep focus ...

next page

Showing results 1-20

Observability Site Reliability Engineer information

What engineer makes $500,000 a year?

A senior or principal Site Reliability Engineer (SRE) or Observability Engineer with extensive experience, specialized skills, and working at large tech companies can earn $500,000 or more annually. Compensation often includes base salary, bonuses, and stock options, especially in high-demand markets and organizations with complex infrastructure.

Is AI replacing SRE?

AI is augmenting the work of Site Reliability Engineers (SREs) by automating tasks such as monitoring, incident detection, and response. However, SREs are still essential for designing systems, managing complex issues, and making strategic decisions that require human judgment. AI tools are considered complementary rather than replacements for SREs' expertise and problem-solving skills.

What engineers make $200,000 a year?

Senior Site Reliability Engineers and Observability Engineers with extensive experience, advanced skills in cloud platforms, automation, and monitoring tools can earn $200,000 or more annually. High compensation often correlates with working at large tech companies, possessing specialized certifications, and managing complex, scalable systems.

What is the difference between Observability Site Reliability Engineer vs Monitoring Engineer?

AspectObservability Site Reliability EngineerMonitoring Engineer
FocusEnsuring system reliability through observability, automation, and incident responseImplementing and managing monitoring tools and dashboards
SkillsCloud platforms, scripting, incident management, observability toolsMonitoring tools, alerting systems, data analysis
Work EnvironmentDevOps teams, cloud infrastructure, large-scale systemsOperations teams, infrastructure monitoring

While both roles involve system health, the Observability Site Reliability Engineer focuses on comprehensive system reliability using observability practices, whereas Monitoring Engineers primarily manage monitoring tools and alerts. The SRE role emphasizes automation, incident response, and system resilience, making it broader in scope.

What engineers make $300,000 a year?

Senior Site Reliability Engineers and Observability Engineers with extensive experience, advanced skills in cloud platforms, automation, and monitoring tools can earn $300,000 or more annually. High compensation often correlates with working at large tech companies, possessing specialized certifications, and taking on leadership or highly technical roles.
What are popular job titles related to Observability Site Reliability Engineer jobs in Iowa? For Observability Site Reliability Engineer jobs in Iowa, the most frequently searched job titles are:
What job categories do people searching Observability Site Reliability Engineer jobs in Iowa look for? The top searched job categories for Observability Site Reliability Engineer jobs in Iowa are:
What cities in Iowa are hiring for Observability Site Reliability Engineer jobs? Cities in Iowa with the most Observability Site Reliability Engineer job openings:
Lead Platform Reliability Engineer

Lead Platform Reliability Engineer

Wells Fargo

West Des Moines, IA • On-site

$100K - $126K/yr

Full-time

Medical, Life, Retirement, PTO

Posted 12 days ago


Wells Fargo rating

7.8

Company rating: 7.8 out of 10

Based on 702 frontline employees who took The Breakroom Quiz

88th of 170 rated banks


Job description

Wells Fargo is seeking a Lead Platform Reliability Engineer to join the CTO Platform organization. This role is designed for highly experienced infrastructure engineers who possess deep technical expertise in one core platform discipline (Network, Middleware, Database, or Storage) and have demonstrated experience collaborating across at least one additional infrastructure domains (Enterprise Tools, Cloud, Observability Tools). The expectation is that this engineer will elevate themselves in looking for trends and patterns that are not limited to these streams and investigate systemic issues that span multiple streams or need a deeper troubleshooting.

As part of our Platform Reliability Engineering (PRE) team, you will apply modern Site Reliability Engineering (SRE) practices to improve the availability, resiliency, observability, scalability, and operational excellence of critical enterprise platforms. You will leverage your domain expertise to identify systemic issues, drive automation, and deliver engineering solutions that strengthen platform stability at scale.

In This Role You Will

  • Serve as the reliability engineering expert for your primary domain (Network, Middleware, Database, or Storage) while partnering across adjacent technology disciplines
  • Lead the investigation and resolution of complex production incidents, identifying root causes and implementing long-term corrective actions
  • Apply SRE principles including service level indicators (SLIs), service level objectives (SLOs), error budgets, and reliability engineering practices to improve platform health
  • Lead capacity analysis, forecasting, and utilization reviews to identify future scaling risks and prevent service degradation before customer impact occurs
  • Perform deep performance analysis across infrastructure layers, identifying bottlenecks, contention points, latency drivers, and resource inefficiencies
  • Identify and remediate configuration drift, operational debt, and platform hygiene issues that impact long-term reliability
  • Drive proactive reliability improvements through observability, automation, performance optimization, and resiliency engineering
  • Design and implement automation solutions that eliminate operational toil, reduce manual intervention, and improve recovery capabilities
  • Define and enhance enterprise observability standards through metrics, logging, tracing, alerting, and service health monitoring
  • Partner closely with engineering, infrastructure, application, cloud, and operations teams to improve platform performance and availability
  • Lead blameless post-incident reviews and convert recurring operational issues into measurable engineering improvements
  • Identify reliability risks and communicate technical recommendations to engineering leaders and senior stakeholders
  • Mentor engineers and technical teams on reliability engineering, operational excellence, automation, and platform best practices

Required Qualifications

  • 5+ years of Systems Engineering, Infrastructure Engineering, Platform Engineering, Technology Architecture, or equivalent experience demonstrated through work experience, military experience, training, or education
  • 5+ years supporting and engineering enterprise-scale production environments
  • 5+ years of experience with hands-on expertise in one of the following technology domains:
    • Network Engineering (routing, switching, load balancing, DNS, network observability, performance analysis)
    • Middleware Engineering (WebSphere, Tomcat, JBoss, Kafka, MQ, application platforms, integration technologies)
    • Database Engineering (Oracle, SQL Server, PostgreSQL, MongoDB, database performance, replication, HA/DR)
    • Storage Engineering (SAN/NAS technologies, storage virtualization, backup/recovery, performance and capacity management)

Desired Qualifications

  • Strong experience applying SRE principles, including SLI/SLO development, error budgets, incident analysis, and reliability measurement
  • Experience supporting highly available, mission-critical production environments
  • Proven success troubleshooting complex issues spanning multiple technology domains in large-scale distributed environments
  • Experience with capacity planning, resiliency engineering, fault tolerance, disaster recovery, and performance optimization
  • Hands-on experience with observability and monitoring platforms such as Grafana, Splunk, Prometheus, AppDynamics, Cribl, ThousandEyes, Dynatrace, or similar technologies
  • Experience building dashboards, alerts, service health indicators, and operational reporting
  • Strong automation and scripting experience using Python, Bash, PowerShell, or similar technologies
  • Experience developing operational tooling, API integrations, self-healing capabilities, and automated remediation solutions
  • Familiarity with Git-based development practices, CI/CD pipelines, infrastructure automation, and Infrastructure as Code tools such as Ansible or Terraform
  • Experience diagnosing and resolving issues that span multiple infrastructure layers
  • Ability to influence technical direction across infrastructure and engineering organizations
  • Experience leading major incident reviews and driving sustainable operational improvements
  • Demonstrated success mentoring engineers and promoting reliability engineering best practices
  • Strong communication skills with the ability to translate technical concepts into business-focused outcomes

Job Expectations:

  • This position offers a hybrid schedule
  • This position does not offer Visa sponsorship

Pay Range

Reflected is the base pay range offered for this position. Pay may vary depending on factors including but not limited to demonstrated examples of prior performance, skills, experience, or work location. Employees may also be eligible for incentive opportunities.

$119,000.00 - $224,000.00

Benefits

Wells Fargo provides eligible employees with a comprehensive set of benefits, many of which are listed below. VisitBenefits - Wells Fargo Jobs for an overview of the following benefit plans and programs offered to employees.

  • Health benefits
  • 401(k) Plan
  • Paid time off
  • Disability benefits
  • Life insurance, critical illness insurance, and accident insurance
  • Parental leave
  • Critical caregiving leave
  • Discounts and savings
  • Commuter benefits
  • Tuition reimbursement
  • Scholarships for dependent children
  • Adoption reimbursement

Posting End Date:

28 Jul 2026

*Job posting may come down early due to volume of applicants.

We Value Equal Opportunity

Wells Fargo is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other legally protected characteristic.

Employees support our focus on building strong customer relationships balanced with a strong risk mitigating and compliance-driven culture which firmly establishes those disciplines as critical to the success of our customers and company. They are accountable for execution of all applicable risk programs (Credit, Market, Financial Crimes, Operational, Regulatory Compliance), which includes effectively following and adhering to applicable Wells Fargo policies and procedures, appropriately fulfilling risk and compliance obligations, timely and effective escalation and remediation of issues, and making sound risk decisions. There is emphasis on proactive monitoring, governance, risk identification and escalation, as well as making sound risk decisions commensurate with the business unit's risk appetite and all risk and compliance program requirements.

Applicants with Disabilities

To request a medical accommodation during the application or interview process, visitDisability Inclusion at Wells Fargo.

Drug and Alcohol Policy

Wells Fargo maintains a drug free workplace. Please see our Drug and Alcohol Policy to learn more.

Wells Fargo Recruitment and Hiring Requirements:

a. Third-Party recordings are prohibited unless authorized by Wells Fargo.

b. Wells Fargo requires you to directly represent your own experiences during the recruiting and hiring process.


What Wells Fargo employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Wells Fargo logo

About Wells Fargo

Sourced by ZipRecruiter

Wells Fargo & Company (NYSE: WFC) is a leading financial services company that has approximately $1.9 trillion in assets, proudly serves one in three U.S. households and more than 10% of small businesses in the U.S., and is a leading middle market banking provider in the U.S. We provide a diversified set of banking, investment and mortgage products and services, as well as consumer and commercial finance, through our four reportable operating segments: Consumer Banking and Lending, Commercial Banking, Corporate and Investment Banking, and Wealth & Investment Management. Wells Fargo ranked No. 41 on Fortune's 2022 rankings of America's largest corporations. In the communities we serve, the company focuses its social impact on building a sustainable, inclusive future for all by supporting housing affordability, small business growth, financial health and a low-carbon economy.

Industry

Finance and insurance

Company size

10,000+ Employees

Headquarters location

San Francisco, CA, US

Year founded

1852

Social media