1

Site Reliability Engineer Intern Jobs in Michigan

Reliability Engineer - Vibration

Midland, MI ยท On-site

$88K - $110K/yr

Coach site reliability engineers and vibration analysts * Develop technical resource networks and ... communities of practice * Mentor engineers, technicians, and condition monitoring specialists ...

Cloud Engineer

Dearborn, MI ยท On-site

$51.25 - $68.50/hr

Apply Site Reliability Engineering (SRE) principles to improve availability, reliability, scalability, and operational performance.Integrate enterprise security, backup, disaster recovery, and data ...

Mechanical Engineering Intern

Hemlock, MI ยท On-site

$15.75 - $21.25/hr

Reliability Engineering Intern - Assist Reliability engineers with supporting our manufacturing ... The candidate must be within driving distance to the site and have reliable transportation ...

Mechanical Engineering Intern

Hemlock, MI

$15.75 - $21.25/hr

Reliability Engineering Intern - Assist Reliability engineers with supporting our manufacturing ... The candidate must be within driving distance to the site and have reliable transportation ...

Reliability Engineer

Battle Creek, MI ยท On-site

$70 - $90/hr

As a Reliability Engineer, you'll be the go-to partner for maintenance and operations, using data ... This role is designed for on-site partnership with maintenance and operations teams, with daily ...

New

Reliability Engineer

Battle Creek, MI ยท On-site

$92K - $116K/yr

As a Reliability Engineer, you'll be the go-to partner for maintenance and operations, using data ... This role is designed for on-site partnership with maintenance and operations teams, with daily ...

Reliability Engineering Intern - Assist Reliability engineers with supporting our manufacturing ... The candidate must be within driving distance to the site and have reliable transportation ...

Reliability Engineering Intern - Assist Reliability engineers with supporting our manufacturing ... The candidate must be within driving distance to the site and have reliable transportation ...

Showing results 21-40

Site Reliability Engineer Intern information

See Michigan salary details

$9

$55

$80

How much do site reliability engineer intern jobs pay per hour?

As of Sep 2, 2026, the average hourly pay for site reliability engineer intern in Michigan is $55.56, according to ZipRecruiter salary data. Most workers in this role earn between $47.79 and $63.46 per hour, depending on experience, location, and employer.

What is a site reliability engineer intern?

A Site Reliability Engineer (SRE) Intern supports the reliability, scalability, and performance of software systems by assisting with automation, monitoring, and incident response. They work closely with development and operations teams to improve system resilience and efficiency. Typical tasks include writing scripts, analyzing system logs, and contributing to documentation or tooling improvements. This role helps interns gain hands-on experience in infrastructure management, cloud services, and DevOps practices.

What types of projects or tasks does a site reliability engineer intern typically work on?

As a Site Reliability Engineer Intern, you may work on projects like automating operational processes, building monitoring dashboards, troubleshooting incidents, and helping improve infrastructure reliability. Interns often collaborate closely with full-time engineers on tasks such as writing scripts to automate deployments, responding to on-call alerts, or optimizing system performance. You'll gain practical experience using industry-standard tools and practices while learning about uptime, scalability, and reliability workflows. It's a great opportunity to develop hands-on skills and understand how large-scale systems are maintained in a professional environment.

What are the key skills and qualifications needed to thrive as a site reliability engineer intern, and why are they important?

To thrive as a Site Reliability Engineer Intern, you should have a solid understanding of computer science fundamentals, programming (such as Python or Go), and basic networking concepts, often supported by relevant coursework or prior technical internships. Familiarity with cloud platforms (like AWS or GCP), containerization tools (e.g., Docker, Kubernetes), and version control systems (Git) is highly valuable. Strong problem-solving abilities, effective communication, and a collaborative mindset help interns stand out in team settings. These skills are crucial for learning quickly, contributing to the team's reliability initiatives, and ensuring highly available and scalable system operations.

What are the most commonly searched types of Site Reliability Engineer jobs in Michigan?

The most popular types of Site Reliability Engineer jobs in Michigan are:

What job categories do people searching Site Reliability Engineer Intern jobs in Michigan look for?

The top searched job categories for Site Reliability Engineer Intern jobs in Michigan are:

What cities in Michigan are hiring for Site Reliability Engineer Intern jobs?

Cities in Michigan with the most Site Reliability Engineer Intern job openings:

Infographic showing various Site Reliability Engineer Intern job openings in Michigan as of August 2026, with employment types broken down into 1% As Needed, 81% Full Time, 12% Part Time, 2% Temporary, and 4% Contract. Highlights an 93% Physical, 3% Hybrid, and 4% Remote job distribution, with an average salary of $115,559 per year, or $55.6 per hour.

Staff Site Reliability Engineer

Sight Machine, Inc.

Ann Arbor, MI โ€ข On-site, Remote

$55.75 - $74/hr

Full-time

Medical, Life, PTO

Posted 15 days ago


Job description

Team Culture
Great things happen when people can bring their authentic selves to work. We empower all of our employees to share their perspectives, passions and experiences because collectively we make a better, stronger team. Our team members collaborate closely with peers & cross functional stakeholders throughout the business, our clients on the forefront of digital transformation, and the cutting edge of digital manufacturing thought leadership.
We take pride in our self-starter culture where employees are enabled and encouraged to achieve their professional goals through leadership guidance, learning and development. Our philosophy is that careers are continuous journeys, and we dedicate time and offer resources so that employees can reach their full potential.
Benefits + Perks
We value you at and outside of work and know your loved ones are important. Our benefits are designed to support you and your family's health through life's expected and unexpected events.
Our Benefits Include:
  • Competitive Salary + Stock Options
  • Health Care Coverage + Life Insurance + Health Savings Account + Flexible Spending
  • Account (includes spouse + children)
  • Flexible Vacation Policy
  • Adaptable Working Schedule and Environment
  • Our Perks Include:
  • Casual Dress Attire
  • Hybrid work flexibility
  • Catered Lunches, Snacks and Beverages
  • Commuter Savings Program
  • Company Outings
  • Designated Volunteering Hours + Group Volunteer Events

Sight Machine is proud to be an equal opportunity employer and considers candidates regardless of age, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. Sight Machine also considers qualified applicants regardless of criminal histories, consistent with legal requirements.
About Sight Machine, Inc.
Sight Machine strengthens manufacturers by providing the industry's only standard data model and system-level visualization capabilities. By integrating all crucial data into a single innovative platform, everyone involved in the fabrication process can visualize, contextualize and examine data in one intuitive interface.
Sight Machine is committed and mission-driven to improve lives, strengthen communities and make the world cleaner through continuously re-envisioning manufacturing processes - making them more efficient, sustainable and absolute.
Founded in Michigan in 2011 and expanded to San Francisco in 2012, Sight Machine blends the spirit of technology innovation and the down to earth style of Detroit manufacturing. Our team includes early leadership from Yahoo, Tesla Motors and Oracle. Together, we share wide industry knowledge and a commitment to advance manufacturing to a more sustainable future.
We take pride in our self-starter culture where employees are enabled and encouraged to achieve their professional goals through leadership guidance, learning and development. Our philosophy is that careers are continuous journeys, and we dedicate time and offer resources so that employees can reach their full potential.
Sight Machine is proud to be an equal opportunity employer and considers candidates legally authorized to work in the US regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. Sight Machine also considers qualified applicants regardless of criminal histories, consistent with legal requirements.
About the team
Sight Machine is built on the shoulders of a unique, robust and highly scalable Infrastructure as Code model. This enables the creation and operation of customer instances in our ecosystem in a standardized and simplified manner. We are looking for team members to help us build, maintain, and improve the infrastructure that makes Sight Machine the leading provider of Manufacturing Data Pipelines and Analytics.
Great things happen when people can bring their authentic selves to work. We empower all of our team members to share their perspectives, passions and experiences because collectively we make a better, stronger team through always "open communications" mind.
Our team collaborates closely with peers & cross functional stakeholders throughout the business, our clients on the forefront of digital transformation, and the cutting edge of digital manufacturing thought leadership.
Sight Machine has offices in San Francisco, CA and Ann Arbor, Mi. We do have a remote-friendly culture with people based all around the US and the rest of the world. For this role in particular, the ideal candidate is located near either of our offices and willing to work in a hybrid capacity. We would still consider 100% remote for exceptional candidates if they aren't located near an office.
About the role
Join the Cloud Infrastructure Team as a technical leader driving reliability, automation, and scalability across the systems running Sight Machine's platform. You'll operate at the intersection of classic SRE discipline which include IaC, CI/CD, observability, incident response and the emerging demands of running agentic AI systems in production: LLM gateways, agent orchestration, and the operational patterns that come with non-deterministic workloads.
This is a senior level IC role. You'll help set and drive technical direction for infrastructure and reliability practices across teams, mentor senior engineers, and be a primary escalation point for the org's hardest systems problems while still being hands-on with code, infrastructure, and incidents.
Success requires deep technical range, sound judgment on risk vs. customer impact, and the ability to influence architecture decisions across Development Engineering without formal authority.
What You'll Actually Work On
  • Champion an agentic-AI-first engineering mindset: identify where AI-driven automation and agent-based tooling can replace manual toil, and hold that work to the same quality, testing, and reliability bar as any other production system
  • Evolve reliability practices for meeting reliability SLO's, error budgets, drive incident postmortems to systemic (not just symptomatic) fixes, and lead reliability reviews for new services before they hit production
  • Troubleshoot and resolve the org's most complex, cross-layer systems problems CI/CD, container orchestration, networking, OS, cloud resources, databases, and increasingly, agentic AI/LLM orchestration layers
  • Design, build, and operate the infrastructure supporting agentic AI workloads, LLM gateway routing, agent orchestration frameworks, monitoring of non-deterministic/AI-driven services, and the operational tooling needed to run them reliably at scale
  • Architect and instrument monitoring, alerting, and observability infrastructure for critical services, with an eye toward what "critical" means for AI-driven systems specifically
  • Author and continuously improve operational runbooks and automation, increasingly incorporating agentic/AI-assisted tooling (e.g., automated triage, AI-assisted incident response) where it measurably reduces toil
  • Design and build internal platforms and developer tooling that other engineers build on top of
  • Participate in on-call coverage and help evolve the program as we scale including escalation paths and reducing avoidable pages through better automation
  • Bring a startup mindset of daily engagement: staying close to what's breaking, what customers are hitting, and where the team needs help, even outside a formal ticket or rotation
  • Mentor senior and mid-level engineers; act as a technical sounding board across teams
  • Proactively identify and drive cross-team initiatives that improve stability, reliability, and availability, this is expected to be self-directed, not assigned

What We're Looking For
  • Demonstrated experience designing, building, or operating agentic AI/LLM-based systems in production, held to the same quality-first, test-driven rigor as traditional infrastructure code, not just prototype-grade work
  • Embody a quality-first and security-first culture in all that you do
  • 10+ years of experience with Kubernetes/Docker in at least one top-tier cloud provider (Azure, GCP, AWS), including production-scale multi-tenant or multi-cluster environments
  • 10+ years coding experience (Python, Go, Java, or similar) with a track record of building tools/platforms used by other engineers, not just scripts
  • 10+ years with IaC and CI/CD tooling (Terraform/OpenTofu, FluxCD or similar GitOps tooling, Jenkins/GitHub Actions)
  • Strong, provable Linux and networking fundamentals (TCP/IP and application-layer)
  • Practical experience integrating or operating LLM/agentic AI systems in a production context this can be API-based orchestration, LLM gateways, or agent frameworks
  • A track record of authoring technical documentation (design docs, ADRs, runbooks) that other engineers actually use
  • Demonstrated mentorship of other engineers, without needing formal management authority to do it
  • Strong bias for action over endless planning, hands-on, has made mistakes, learned from them, and can weigh risk vs. customer impact under pressure
  • Clear, empathetic communicator, comfortable pushing back on architecture decisions across teams
  • Operational experience with monitoring/alerting systems (Prometheus, Grafana, Loki, Sentry, Signoz or equivalents)
  • Deep understanding of cloud performance, able to diagnose and resolve bottlenecks others can't

Nice to Have
  • Experience with elements of our current tech stack are a plus: Kubernetes, FluxCD, Terraform, Helm Charts, Prometheus, Elasticsearch, Python, Java, Kafka, Postgres, and Jenkins
  • Previous experience or a keen interest in industrial IoT, analytics, or manufacturing a plus