1

Site Reliability Manager Jobs in Seattle, WA (NOW HIRING)

Manager, Site Reliability Engineering San Francisco, California Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted ...

SRE Engineer

Redmond, WA · On-site

$63.75 - $84.75/hr

Deploy and manage AI resources on Microsoft Azure, including AI Foundry and RAG solutions * Monitor and ensure service uptime, availability, reliability, and latency * Track and integrate SRE metrics ...

Sr. SRE Consultant

Seattle, WA · On-site

$64.75 - $86.25/hr

Role: Sr. SRE (Very Strong Technical SRE) Location: Seattle, WA WFO: Mandatory (3 days/week) Short ... Experience in ITSM process including Incident, Problem, and Change management. * Experience in ...

SRE Engineer -AI

Redmond, WA · On-site

$63.75 - $84.75/hr

Deploy and manage AI resources on Microsoft Azure, including AI Foundry and RAG solutions Monitor and ensure service uptime, availability, reliability, and latency Track and integrate SRE metrics ...

The Apple Service Engineering - Data Streaming SRE team is looking for Site Reliability Engineers with experience developing processes, tools, and automation for managing distributed systems in ...

Site Reliability Engineer

Seattle, WA · On-site

$64.75 - $86.25/hr

We offer IT solutions across the disciplines of program/project management, applications ... Site Reliability Engineer Duration: Long Term These are top 4 criteria important for SREs 2. ...

The Apple Service Engineering - Data Streaming SRE team is looking for Site Reliability Engineers with experience developing processes, tools, and automation for managing distributed systems in ...

Site Reliability Engineer

Redmond, WA · On-site

$63.75 - $84.75/hr

REDMOND, Washington (Remote) working PT hours Role Overview This is a site reliability engineering ... Manage assigned projects and program components to deliver services in accordance with established ...

Senior Site Reliability Engineer

Bellevue, WA

$64.25 - $85.50/hr

The Senior Site Reliability Engineer Opportunity Reporting to the Manager, Site Reliability Engineering , this role will help build, improve, and maintain our cloud platform services by designing and ...

The Senior Site Reliability Engineer Opportunity Reporting to the Manager, Site Reliability Engineering , this role will help build, improve, and maintain our cloud platform services by designing and ...

Senior Site Reliability Engineer

Bellevue, WA · On-site

$64.25 - $85.50/hr

The Senior Site Reliability Engineer Opportunity Reporting to the Manager, Site Reliability Engineering , this role will help build, improve, and maintain our cloud platform services by designing and ...

This SRE will configure, tune, and fix multi-tiered systems to achieve optimal application ... We manage jobs as well as applications on bare-metal and cloud computing platforms to deliver data ...

next page

Showing results 1-20

Site Reliability Manager information

See Seattle, WA salary details

$70.6K

$133.7K

$191.8K

How much do site reliability manager jobs pay per year?

As of Aug 25, 2026, the average yearly pay for site reliability manager in Seattle, WA is $133,704.00, according to ZipRecruiter salary data. Most workers in this role earn between $107,500.00 and $159,300.00 per year, depending on experience, location, and employer.

What is the difference between Site Reliability Manager vs DevOps Engineer?

AspectSite Reliability ManagerDevOps Engineer
CredentialsTypically requires a Bachelor's in Computer Science, certifications like SRE or Cloud certificationsOften holds a Bachelor's in Computer Science or related field, with certifications in cloud platforms or automation tools
Work EnvironmentLeads teams managing large-scale systems, focusing on reliability and uptimeWorks on automation, CI/CD pipelines, and infrastructure deployment
Industry UsageCommon in tech, cloud services, and large enterprisesWidely used in startups, tech companies, and organizations adopting DevOps practices

The main difference is that Site Reliability Managers focus on ensuring system reliability and managing SRE teams, while DevOps Engineers concentrate on automation, deployment, and continuous integration. Both roles require technical expertise but serve different strategic objectives within IT operations.

What is the role of a site reliability manager?

A site reliability manager (SRM) is responsible for ensuring the reliability, availability, and performance of a company's IT systems and services. They often oversee incident response, implement automation tools, and collaborate with development teams to improve system stability and scalability, typically requiring knowledge of monitoring tools and scripting skills.

Senior Site Reliability Engineer (SRE)

Socket.dev

Bellevue, WA • On-site

$92.25 - $140/hr

Other

Medical, Dental, Vision, Retirement, PTO

Posted 6 days ago


Job description

Title: Senior Site Reliability Engineer (SRE)

Location: Bellevue, WA - Hybrid (3-4 days onsite per week)

As a Senior Site Reliability Engineer, you will help build, scale, and operate the resilient platforms that power critical digital experiences across web, mobile, API, AI/ML, and customer-facing environments. This role is ideal for an engineer who thrives at the intersection of software, cloud infrastructure, automation, and operations, and who is passionate about improving reliability, observability, scalability, and developer experience.

You will collaborate closely with software engineering, platform, architecture, product, and security teams to strengthen platform performance and availability, reduce operational toil, and advance intelligent, automated operations across mission‑critical services.

Responsibilities
  • Design, implement, and support highly available, scalable, and resilient platform services.
  • Define and manage Service Level Objectives (SLOs), Service Level Indicators (SLIs), and error budgets.
  • Identify and address reliability risks and performance bottlenecks across distributed systems.
  • Lead root cause analysis for critical incidents and drive sustainable corrective actions.
  • Participate in production support and on‑call rotations for business‑critical applications.
  • Build and enhance observability solutions using tools such as Splunk, Grafana, Prometheus, Datadog, New Relic, and OpenTelemetry.
  • Create actionable dashboards, alerts, metrics, and reporting that improve operational visibility.
  • Drive continuous improvement in Mean Time to Detect (MTTD) and Mean Time to Resolution (MTTR).
  • Support and evolve shared platform capabilities used across multiple engineering teams.
  • Develop self‑service platform features, reusable services, and automation frameworks that improve developer productivity.
  • Partner with architecture and engineering teams to define future‑state platform strategies.
  • Design and manage cloud‑native infrastructure across AWS and Azure environments.
  • Implement Infrastructure as Code using Terraform, CloudFormation, Helm, and Kubernetes manifests.
  • Automate provisioning, deployment, configuration management, and recovery processes.
  • Design and optimize CI/CD pipelines to enable secure, reliable, and efficient software delivery.
  • Improve deployment speed and quality through automation, release validation, and deployment controls.
  • Support GitOps operating models and deployment automation practices.
  • Lead operational readiness reviews, disaster recovery exercises, and resiliency initiatives.
  • Establish runbooks, playbooks, and automated remediation solutions.
  • Drive chaos engineering and resiliency testing efforts.
  • Ensure alignment with enterprise operational standards and security requirements.
  • Leverage AI‑assisted operations, observability, and incident intelligence capabilities to proactively identify and mitigate risk.
  • Advance intelligent platform capabilities that enhance engineering efficiency and operational excellence.
Qualifications
  • Bachelor’s degree in Computer Science, Engineering, Information Technology, or a related field, or equivalent practical experience.
  • 5+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Operations.
  • Experience supporting large‑scale distributed applications in production environments.
  • Experience operating customer‑facing digital platforms with high availability expectations.
  • Strong Linux administration and troubleshooting skills.
  • Expertise with Kubernetes and containerized environments.
  • Experience with AWS and/or Azure cloud platforms.
  • Strong scripting and automation skills in Python and Bash; Go is preferred.
  • Strong experience with Terraform, Git, CI/CD pipelines, Docker, and Kubernetes.
  • Experience working with APIs, microservices, and distributed architectures.
  • Hands‑on experience with observability and monitoring tools such as Splunk, Grafana, Prometheus, OpenTelemetry, New Relic, and cloud‑native monitoring platforms.
  • Demonstrated strength in incident management, escalation leadership, root cause analysis, and problem management.
  • Experience with capacity planning, performance engineering, disaster recovery, and resiliency testing.
  • Preferred: experience supporting mobile applications (iOS and Android) and digital customer platforms.
  • Preferred: experience with API gateways, CDN technologies, and edge architectures.
  • Preferred: knowledge of AI/ML platform operations.
  • Preferred: familiarity with cybersecurity best practices and DevSecOps.
  • Preferred: AWS Certified DevOps Engineer, Solutions Architect, Kubernetes, or related certifications.
  • Preferred: experience working within Agile and Product operating models.

The base salary range for this position is $92,250 - $140,000 plus incentives that align with individual and company performance. Actual salaries will vary based on work location, qualifications, skills, education, experience, and competencies. Benefits available to eligible employees in this role include medical, dental, and vision insurance, comprehensive employee assistance program, 401(k) retirement plan, paid time off and holidays.

Physical and Mental Requirements: The employee is regularly required to operate a computer, keyboard, telephone/headset, and/or other office equipment as essential functions of this position. Work is generally sedentary in nature.

Equal Employment Opportunity

Concentrix is an equal opportunity and affirmative action (EEO-AA) employer. We promote equal opportunity to all qualified individuals and do not discriminate in any phase of the employment process based on race, color, religion, sex, sexual orientation, gender identity, national origin, age, pregnancy or related condition, disability, status as a protected veteran, or any other basis protected by law.

  • English (https://www.eeoc.gov/sites/default/files/2023-06/22-088_EEOC_KnowYourRights6.12.pdf)
  • Spanish (https://www.eeoc.gov/sites/default/files/2023-06/22-088_EEOC_KnowYourRightsSp6.12.pdf)
#J-18808-Ljbffr