1

Site Reliability Engineering Jobs in Michigan (NOW HIRING)

Reliability Engineer - Vibration

Midland, MI · On-site

$88K - $110K/yr

Experience leading enterprise or multi-site improvement initiatives * Strong communication and influence skills Preferred: * ISO Category III or IV Vibration Certification * Reliability Engineering ...

Cloud Engineer

Dearborn, MI · On-site

$51.25 - $68.50/hr

Apply Site Reliability Engineering (SRE) principles to improve availability, reliability, scalability, and operational performance.Integrate enterprise security, backup, disaster recovery, and data ...

Cloud Engineer

Dearborn, MI · On-site

$51.25 - $68.50/hr

Apply Site Reliability Engineering (SRE) principles to improve system resilience. * Drive root cause analysis and permanent corrective actions. Required Qualifications Education * Bachelor's or ...

Champion modern engineering practices including automation, observability, DevOps, Site Reliability Engineering (SRE), and AI Operations (AIOps). Required Experience, Education and Skills * 3+ years ...

... site reliability engineering * Proficiency with container orchestration (Nomad, Kubernetes), IaaC, and Linux * Strong scripting and programming skills (Python, Bash, Go) * Experience with hybrid ...

Senior Systems Engineer

Southfield, MI · Hybrid

$95K - $131K/yr

Site Reliability Engineering (SRE) experience. * Experience with Active Directory and file systems management. * Experience with TLS, digital certificates, and related troubleshooting. * Familiarity ...

Showing results 21-40

Site Reliability Engineering information

See Michigan salary details

$9

$55

$80

How much do site reliability engineering jobs pay per hour?

As of Sep 5, 2026, the average hourly pay for site reliability engineering in Michigan is $55.56, according to ZipRecruiter salary data. Most workers in this role earn between $47.79 and $63.46 per hour, depending on experience, location, and employer.

What is site reliability engineering?

Site Reliability Engineering (SRE) is a discipline that combines aspects of software engineering and IT operations to build and run scalable, reliable, and efficient systems. SRE teams are responsible for ensuring the availability, performance, and reliability of critical services by automating manual processes, monitoring systems, and responding to incidents. They often work closely with development teams to improve system architecture, deploy new features safely, and maintain service-level objectives (SLOs). Overall, SRE aims to create a bridge between development and operations to deliver robust and reliable software services.

How does a site reliability engineer typically collaborate with development and operations teams?

Site Reliability Engineers (SREs) work closely with both development and operations teams to ensure systems are reliable, scalable, and efficient. They often participate in code reviews, help define service level objectives (SLOs), and develop automation tools to streamline deployment and incident response. SREs act as a bridge between development and IT, translating operational needs into engineering solutions and vice versa. Regular communication and joint problem-solving are essential parts of the role, fostering a culture of shared responsibility for system uptime and performance.

What are the key skills and qualifications needed to thrive as a site reliability engineer, and why are they important?

To thrive as a Site Reliability Engineer, you need a solid background in software engineering, systems administration, and troubleshooting, often supported by a degree in computer science or related field. Familiarity with automation tools, cloud platforms (such as AWS, GCP, or Azure), containerization (Docker, Kubernetes), and monitoring systems is typically required. Strong problem-solving skills, effective communication, and a proactive mindset set outstanding SREs apart. These skills ensure high system reliability, efficient incident response, and seamless collaboration across development and operations teams.

What are the most commonly searched types of Site Reliability Engineering jobs in Michigan?

The most popular types of Site Reliability Engineering jobs in Michigan are:

Infographic showing various Site Reliability Engineering job openings in Michigan as of August 2026, with employment types broken down into 1% As Needed, 77% Full Time, 20% Part Time, and 2% Contract. Highlights an 93% Physical, 3% Hybrid, and 4% Remote job distribution, with an average salary of $115,559 per year, or $55.6 per hour.

Platform Engineer - Infrastructure

Ford Motor Company

Dearborn, MI • On-site

$52.50 - $69.75/hr

Full-time

Medical, Dental, Vision, Life, PTO

Re-posted 3 days ago


Ford Motor Company rating

7.6

Company rating: 7.6 out of 10

Based on 527 frontline employees who took The Breakroom Quiz

11th of 45 rated automakers


Job description

We made history and now we work to transform the future - for our customers, our communities and our families. You'll see your work on the road every day, helping people move freely and pursue their dreams. At Ford, you can build more than vehicles. Come build what matters.

The Ford Motor Credit Company team helps put people behind the wheels of great Ford and Lincoln vehicles. By partnering with dealerships, we provide financing, personalized service and professional expertise to thousands of dealers and millions of customers in over one hundred countries around the world.

In this position...

Ford Credit Platform Engineering builds and operates the shared infrastructure and paved paths that help product teams deliver securely, reliably, and quickly. As a Cloud Infrastructure & SRE Engineer on our Platform Engineering team, you will design, build, and operate the cloud platforms and reliability practices that hundreds of engineers depend on every day.

This role leans toward cloud infrastructure, DevOps, and Site Reliability Engineering, with strong software development skills. You will write code to automate infrastructure, define reliability targets, and create self-service workflows that eliminate toil. You will operate what you build, participate in on-call rotations, and drive systemic improvements from every incident.

About the Team

Our Platform Engineering team sits at the intersection of product and infrastructure. We treat our platform capabilities as products, with users, documentation, support, and a roadmap driven by real feedback. We partner closely with product engineering teams, security, and SRE to standardize patterns around identity, networking, CI/CD, secrets management, and deployment.

We believe the best platform work is invisible. Developers should not have to think about cluster wiring, pipeline configuration, or operational boilerplate. They should focus on their features, and the platform should just work.

How We Work

  • Automate first: Eliminate repeatable manual work. Measure and reduce toil.
  • Reliability is a feature: Design for failure with timeouts, retries with jitter, idempotency, and graceful degradation.
  • Small, safe changes: Incremental delivery, clear rollback strategies, and continuous improvement.
  • Engineering excellence: Design reviews, blameless postmortems, and strong documentation and runbooks.

What Success Looks Like

  • Platform capabilities are easy to adopt, well-documented, and measurably reduce lead time for change.
  • Reliability improves over time, measured by SLO attainment, reduced incident frequency and severity, and faster MTTR.
  • Security posture improves through secure-by-default patterns and automated controls.

You'll have...

  • Bachelors of Science Degree
  • 7 year's experience in software development
  • Experience operating production cloud platforms and services (e.g., GCP/AWS/Azure) with an SRE mindset.
  • Strong fundamentals in Linux, networking, distributed systems, and debugging complex production issues.
  • Proficiency with infrastructure as code and automation (e.g., Terraform, Helm/Kustomize, GitOps tooling).
  • Experience with containers and orchestration (Docker, Kubernetes) and modern CI/CD.
  • Programming and scripting ability (e.g., Go, Python, Java, TypeScript) to build tooling and automate workflows.
  • Clear communication, effective incident leadership, and a customer-focused approach to platform work.

Even better, you may have...

  • Experience defining SLIs/SLOs and implementing SLO-based alerting and dashboards.
  • Observability platform experience (e.g., Prometheus/Grafana, OpenTelemetry, centralized logging).
  • Policy-as-code and supply chain security (e.g., OPA/Rego, SLSA concepts, SBOMs, artifact signing).
  • Experience building golden paths (container images, templates, reference architectures, paved pipelines) adopted by multiple teams.
  • Cost optimization experience (FinOps practices, capacity forecasting, right-sizing, multi-tenant platform controls).

You may not check every box, or your experience may look a little different from what we've outlined, but if you think you can bring value to Ford Motor Company, we encourage you to apply!

As an established global company, we offer the benefit of choice. You can choose what your Ford future will look like: will your story span the globe, or keep you close to home? Will your career be a deep dive into what you love, or a series of new teams and new skills? Will you be a leader, a changemaker, a technical expert, a culture builder...or all of the above? No matter what you choose, we offer a work life that works for you, including:

Immediate medical, dental, vision and prescription drug coverage

Flexible family care days, paid parental leave, new parent ramp-up programs, subsidized back-up child care and more

Family building benefits including adoption and surrogacy expense reimbursement, fertility treatments, and more

Vehicle discount program for employees and family members and management leases

Tuition assistance

Established and active employee resource groups

Paid time off for individual and team community service

A generous schedule of paid holidays, including the week between Christmas and New Year's Day

Paid time off and the option to purchase additional vacation time.

This position is a salary grade 6 and ranges from $85,400-$143,200. 

This position is a salary grade 7 and ranges from $99,600-$166,600. 

This position is a salary grade 8 and ranges from $115,000-$192,900.   

Final determination of salary grade will be based on candidate's skills and experience, and base salary will be set within the applicable range according to job scope, responsibility and competitive market value.

For more information on salary and benefits, click here: https://fordcareers.co/GSR

Visa sponsorship is not available for this position.

Candidates for positions with Ford Motor Company must be legally authorized to work in the United States. Verification of employment eligibility will be required at the time of hire.

We are an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, religion, color, age, sex, national origin, sexual orientation, gender identity, disability status or protected veteran status. In the United States, if you need a reasonable accommodation for the online application process due to a disability, please call 1-888-336-0660.

This position is hybrid. Candidates who are in commuting distance to a Ford hub location may be required to be onsite four or more days per week. 

#LI-Hybrid     #LI-FordCredit    #LI-AF1

What you'll do...

  • Design, build, and operate cloud infrastructure and platform capabilities (networking, compute, Kubernetes, CI/CD, secrets, certificates, identity).
  • Define and improve reliability using service-level indicators (SLIs), service-level objectives (SLOs), and error budgets.
  • Implement observability (metrics, logs, traces) with actionable alerting focused on user impact.
  • Create self-service workflows and automation (infrastructure as code, GitOps, build/release pipelines) that reduce toil.
  • Improve security and compliance through least-privilege access, secure defaults, policy-as-code, and continuous hardening.
  • Participate in on-call rotation, incident response, and post-incident reviews; drive systemic fixes and runbook quality.
  • Partner with application teams to improve deployability, resilience, and cost efficiency (capacity planning, autoscaling, graceful degradation).

What Ford Motor Company employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom