1

Site Reliability Engineer Sre Jobs in Michigan (NOW HIRING)

Site Reliability Engineer

Rochester, MI ยท On-site

$52.50 - $69.75/hr

Description Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 ...

Site Reliability Engineer

Birmingham, MI ยท On-site

$54.25 - $72.25/hr

Description Site Reliability Engineer Location: Remote, United States Employment Type: Full-Time Benefits Offered: Vision, Medical, Life, Dental, 401K Gross Annual Base Salary: USD 114,000-148,000 ...

Site Reliability Engineer

Auburn Hills, MI ยท On-site

$54 - $71.75/hr

Required Qualifications * 5+ years of experience in monitoring, observability, or SRE roles. * 3+ years of hands-on experience with: * Splunk (Search Processing Language - SPL) * AppDynamics ...

Site Reliability Engineer

Dearborn, MI

$52.25 - $69.50/hr

Site Reliability Engineer #1063726 Position Description: Employees in this job function are responsible for ensuring availability, reliability and performance of cloud and network systems and ...

New

Site Reliability Engineer

Auburn Hills, MI ยท On-site

$54 - $71.75/hr

Required Qualifications * 5+ years of experience in monitoring, observability, or SRE roles. * 3+ years of hands-on experience with: * Splunk (Search Processing Language - SPL) * AppDynamics ...

SRE Engineer

Dearborn, MI ยท On-site

$52.75 - $70/hr

As a Site Reliability Engineer at Ford Motor Company, you will play a pivotal role in elevating the performance and dependability of our Marketing and Sales Tech platform and applications. In this ...

Site Reliability Engineer II

Detroit, MI ยท On-site

$56.50 - $75/hr

Site Reliability Engineer II The SRE II sits at the intersection of software engineering and platform operations. You will own the reliability, scalability, and operational hygiene of Kastle's core ...

Site Reliability Engineer III

Ann Arbor, MI ยท On-site

$55.75 - $74/hr

Level III Site Reliability Engineers are recognized technical experts who lead complex projects and initiatives, drive innovation, and serve as key resources for both their team and the broader ...

Site Reliability Engineer III

Ann Arbor, MI ยท On-site

$55.75 - $74/hr

Level III Site Reliability Engineers are recognized technical experts who lead complex projects and initiatives, drive innovation, and serve as key resources for both their team and the broader ...

As a Site Reliability Engineer, you will be focused on supporting a specific, highly available, and very secure production environment. You will analyze the reliability of our environments, use your ...

You'll operate at the intersection of classic SRE discipline which include IaC, CI/CD, observability, incident response and the emerging demands of running agentic AI systems in production: LLM ...

next page

Showing results 1-20

Site Reliability Engineer Sre information

What is a site reliability engineer (SRE)?

A Site Reliability Engineer (SRE) is a professional who combines software engineering and IT operations skills to ensure reliable, scalable, and efficient systems. SREs automate system operations, monitor system health, and manage incident response to minimize downtime and ensure high availability. Their work bridges the gap between development and operations teams, focusing on building robust infrastructure and processes. SREs often use tools and practices like automation, monitoring, and performance tuning to maintain service reliability. They also set and measure service level objectives (SLOs) to align system performance with business goals.

What are the key skills and qualifications needed to thrive as a site reliability engineer (SRE), and why are they important?

To thrive as a Site Reliability Engineer (SRE), you need a solid background in software engineering, systems administration, and automation, often supported by a degree in computer science or a related field. Familiarity with cloud platforms (like AWS, GCP, or Azure), containerization (Docker, Kubernetes), and monitoring tools (Prometheus, Grafana), as well as scripting languages such as Python or Bash, is typically required. Strong problem-solving, communication, and collaboration skills help SREs manage incidents and work effectively with development and operations teams. These competencies are essential to ensure system reliability, optimize performance, and maintain seamless service availability.

What are some of the most common challenges faced by site reliability engineers (SREs), and how can they be addressed?

Site Reliability Engineers often face challenges such as managing the balance between reliability and rapid feature deployment, handling on-call responsibilities, and automating manual processes. To address these, SREs work closely with development and operations teams to implement strong monitoring, establish clear service level objectives (SLOs), and continually improve incident response procedures. Building a culture of blameless postmortems and investing in automation can also help reduce repetitive work and improve overall system reliability.

What is the difference between Site Reliability Engineer Sre vs DevOps Engineer?

AspectSite Reliability Engineer SreDevOps Engineer
CredentialsTypically requires experience in systems engineering, scripting, and monitoring toolsOften has certifications in cloud platforms, automation, and CI/CD tools
Work EnvironmentFocuses on maintaining system reliability, scalability, and incident responseEmphasizes automation, deployment, and continuous integration/delivery
Industry UsageCommon in tech, finance, and large-scale cloud servicesWidely used across startups, tech companies, and enterprises

While both roles aim to improve system performance and automation, Site Reliability Engineers Sre primarily focus on reliability and incident management, whereas DevOps Engineers concentrate on deployment automation and continuous integration. The roles often overlap but serve distinct core functions within IT and software development teams.

What job categories do people searching Site Reliability Engineer Sre jobs in Michigan look for?

The top searched job categories for Site Reliability Engineer Sre jobs in Michigan are:

What cities in Michigan are hiring for Site Reliability Engineer Sre jobs?

Cities in Michigan with the most Site Reliability Engineer Sre job openings:

Infographic showing various Site Reliability Engineer Sre job openings in Michigan as of August 2026, with employment types broken down into 87% Full Time, and 13% Contract. Highlights an 70% In-person, 17% Hybrid, and 13% Remote job distribution.

Site Reliability Engineer

OneStream Software

Rochester, MI โ€ข On-site

$52.50 - $69.75/hr

Other

Medical, Dental, Vision, Life, Retirement, PTO

Re-posted 29 days ago


Key responsibilities

  • Implement application/infrastructure observability solutions to ensure application availability, reliability, and performance.

  • Participate in regular on-call rotations and communicate incident resolutions through post-mortem reports and review meetings.

  • Proactively collaborate with Product and Engineering teams to develop, deploy, and maintain reliable systems and services.


Job description

Description

Site Reliability Engineer


Location:
Remote, United States

Employment Type: Full-Time

Benefits Offered: Vision, Medical, Life, Dental, 401K

Gross Annual Base Salary: USD 114,000-148,000
Additional variable compensation and benefits may apply. Total compensation is based on experience, skills, and location using objective, job-related criteria.

Summary

As a Site Reliability Engineer, you will focus on ensuring the platform and services customers rely on are reliable, performant, and highly available. If you enjoy staying at the forefront of technology and automating infrastructure deployments, then this is the job for you. This vital role within Cloud Services requires knowledge and experience designing, implementing, and monitoring scalable and secure cloud services. The employee is expected to work well in a small team and willing to share responsibilities with other team members as needed. You will interact with internal staff, managers, and customers to implement and maintain operations. A passion for technology and learning, and the ability to grow others are vital for success in this role.

Primary Duties and Responsibilities

  • Implement application/infrastructure observability solutions to ensure desired application availability, reliability, and performance.
  • Participate in regular On-Call rotations and share details related to incidents and their resolution through post-mortem reports and regular review meetings.
  • Proactively partner with Product and Engineering teams to identify, develop, deploy, and maintain reliable systems and services.
  • Influence and create new designs, architectures, standards, and methods for large-scale systems.
  • Sustain a high level of reliability for key services and automated systems.
  • Automate processes to improve reliability, performance, and availability.
  • Update technical documentation, workflows, and knowledge base articles.
  • Provide feedback in pull requests and peer coding reviews.
  • Implement codified automated solutions that build integrations between Dynatrace, Azure DevOps and Jira.
  • Solid knowledge in focused areas of OneStream Software.
  • Ability to mentor others in several technical areas.
  • Understanding practical use of SOC/FedRAMP controls to assist Compliance and Security teams.

Required Education and Experience

  • BS/BA in computer science, engineering, or technology-related field (or equivalent work experience).
  • Proven work experience as a Site Reliability Engineer or in a similar role.
  • 6+ years of cloud infrastructure and software development experience.
  • 2+ years hands on experience of Azure Kubernetes Services (AKS) with container-based deployment skills or other platforms such as OpenShift, GKS, EKS.
  • Advanced understanding of APM and observability tools such as Dynatrace, AppInsights, DataDog, Log Analytics, New Relic, Prometheus and Grafana.
  • Advanced understanding of Infrastructure-as-Code (IaC) concepts and tooling (Terraform, CloudFormation templates, Bicep or ARM templates) on Microsoft Azure, Amazon Web Services (AWS), or Google Cloud Platform (GCP).
  • Deep knowledge of Configuration Management/Orchestration utilities such as Ansible, PowerShell DSC, Chef, and Puppet.
  • Advanced understanding of cloud concepts including elasticity, security, and identity management.
  • Well versed familiarity with Agile Development methodologies utilizing Jira or Azure DevOps Boards.
  • 6+ years of hands-on experience with the following technologies, tools, and concepts:
    • Automating processes using PowerShell, Bash, CLI, REST APIs, python, ARM Templates or other scripting languages.
    • Comfortable leveraging source control tools such as Git, Azure DevOps, or GitHub.
    • Knowledge of container orchestration platforms such as Kubernetes, OpenShift, AKS, GKS or helm.
    • Microsoft Azure, Amazon Web Services (AWS) or Google Cloud (GCP).

Preferred Education and Experience

  • Experience working for a cloud service provider (CSP), managed service provider (MSP), or SaaS provider.
  • 6+ years of relevant Azure experience deploying and managing leveraging Infrastructure-as-Code (IAC) concepts.
  • Experience with Microsoft and .NET (.NET, C#, SQL).
  • Experience writing efficient and reliable code in a development environment.
  • Debian, Ubuntu, Alpine or other distributions of the Linux operating systems.
  • Deep knowledge and understanding of containerized applications, with special attention to reliability and monitoring of those containerized applications.

Knowledge, Skills, and Abilities

  • Deal well with ambiguous/undefined problems.
  • Ability to self-motivate and work independently.
  • Strong organizational and prioritization skills.
  • Ability to find and apply effective solutions to emerging problems and challenges.
  • Strong attention to detail.
  • Comfortable communicating with all levels of management and engineering.
  • Ability to get up to speed quickly with modern technologies and services.
  • Ability to multitask on a variety of projects.


Travel

  • Travel Requirement: Travel is not expected to exceed 5%.


Who We Are

OneStream is how today's Finance teams can go beyond just reporting on the past and Take Finance Further by steering the business to the future. It's the only enterprise finance platform that unifies financial and operational data, embeds AI for better decisions and productivity, and empowers the CFO to become a critical driver of business strategy and execution. Our vision is to be the AI operating system for modern finance, digitizing core financial functions and empowering the CFO to become a critical driver of business strategy. To learn more visit www.onestream.com.


Why Join The OneStream Team

  • Transparency around corporate structure, salary, and benefits.
  • Core value of customer success.
  • Variety of project work (not industry-specific).
  • Strong culture andcamaraderie.
  • Multiple training opportunities.


Benefits at OneStream

OneStream employees are passionate, hardworking individuals who go above and beyond to keep our customers happy and follow through on our mission statement. They consistently deliver the best and in turn, we make every effort to keep them cared for and happy. A sample of the benefits we provide are:

  • Excellent Medical Plan.
  • Dental & Vision Insurance.
  • Life Insurance.
  • Short & Long Term Disability.
  • Vacation Time.
  • Paid Holidays.
  • Professional Development.
  • Retirement Plan.

All candidates must be legally authorized to work for any company in the country where this position is located without sponsorship.

OneStream is an Equal Opportunity Employer.


#LI-AS1
#LI-Remote


Equal Opportunity Employer/Protected Veterans/Individuals with Disabilities
This employer is required to notify all applicants of their rights pursuant to federal employment laws. For further information, please review the Know Your Rights notice from the Department of Labor.