1

Freelance Site Reliability Engineer Jobs in Michigan

You'll operate at the intersection of classic SRE discipline which include IaC, CI/CD, observability, incident response and the emerging demands of running agentic AI systems in production: LLM ...

Reliability Engineer - Vibration

Midland, MI ยท On-site

$88K - $110K/yr

Coach site reliability engineers and vibration analysts * Develop technical resource networks and ... communities of practice * Mentor engineers, technicians, and condition monitoring specialists ...

AWS SRE Consultant Summary Vertical Relevance is looking for an AWS SRE Consultant, to join our team as a full-time employee in our New York or New Jersey office or work remotely. This person is ...

New

Cloud Engineer

Dearborn, MI ยท On-site

$51.25 - $68.50/hr

Apply Site Reliability Engineering (SRE) principles to improve availability, reliability, scalability, and operational performance.Integrate enterprise security, backup, disaster recovery, and data ...

Reliability Engineer

Battle Creek, MI ยท On-site

$92K - $116K/yr

As a Reliability Engineer, you'll be the go-to partner for maintenance and operations, using data ... This role is designed for on-site partnership with maintenance and operations teams, with daily ...

Reliability Engineer

Battle Creek, MI ยท On-site

$80 - $100/hr

As a Reliability Engineer, you'll be the go-to partner for maintenance and operations, using data ... This role is designed for on-site partnership with maintenance and operations teams, with daily ...

Showing results 21-40

Freelance Site Reliability Engineer information

What is a freelance site reliability engineer?

Freelance Site Reliability Engineers (SREs) are independent professionals who help organizations maintain the reliability, scalability, and performance of their software systems. They combine software engineering and IT operations skills to automate processes, monitor systems, and respond to incidents. Working on a contract or project basis, freelance SREs often collaborate with multiple clients to ensure services run smoothly and efficiently. They may be responsible for tasks like infrastructure as code, incident response, monitoring, and improving system resilience. Their flexible engagement model allows organizations to access specialized expertise without committing to full-time hires.

How does a freelance site reliability engineer typically collaborate with client teams to ensure reliable service delivery?

As a Freelance Site Reliability Engineer, you often work closely with client development and operations teams to understand their infrastructure, incident response protocols, and deployment pipelines. Frequent communication through virtual meetings, shared documentation, and messaging platforms is essential to align on service-level objectives and address reliability concerns. You may be responsible for proactively identifying potential issues, recommending improvements, and sometimes being on-call for critical incidents. Building trust and integrating smoothly with remote teams is key, as you'll often need to adapt to diverse workflows and technical stacks.

What are the key skills and qualifications needed to thrive as a freelance site reliability engineer, and why are they important?

To thrive as a Freelance Site Reliability Engineer, you need a deep understanding of systems administration, cloud infrastructure, coding in languages like Python or Go, and a proven track record in managing distributed systems. Familiarity with tools such as Kubernetes, Docker, CI/CD pipelines, and monitoring platforms like Prometheus or Datadog, as well as certifications (e.g., AWS Certified Solutions Architect), is highly beneficial. Strong problem-solving, communication, and time management skills help you excel in client-driven, fast-paced environments. These capabilities ensure high system availability, rapid incident resolution, and effective stakeholder collaboration for reliable service delivery.

What is the difference between Freelance Site Reliability Engineer vs Freelance DevOps Engineer?

AspectFreelance Site Reliability EngineerFreelance DevOps Engineer
CredentialsRelevant certifications (e.g., SRE, cloud certifications)DevOps certifications, cloud expertise
Work EnvironmentFocus on system reliability, uptime, and incident responseFocus on deployment, automation, and CI/CD pipelines
Industry UsageTech companies, cloud providers, SaaS firmsSoftware development, IT services, startups
Search & Comparison IntentUnderstanding reliability roles in freelance workComparing automation and deployment roles

Freelance Site Reliability Engineers primarily focus on maintaining system uptime, incident management, and reliability metrics, while Freelance DevOps Engineers concentrate on automation, deployment pipelines, and continuous integration. Both roles require cloud and scripting skills, but SREs emphasize system stability and incident response, whereas DevOps roles focus on deployment efficiency and automation processes.

What are the most commonly searched types of Site Reliability Engineer jobs in Michigan?

The most popular types of Site Reliability Engineer jobs in Michigan are:

What are popular job titles related to Freelance Site Reliability Engineer jobs in Michigan?

For Freelance Site Reliability Engineer jobs in Michigan, the most frequently searched job titles are:

What job categories do people searching Freelance Site Reliability Engineer jobs in Michigan look for?

The top searched job categories for Freelance Site Reliability Engineer jobs in Michigan are:

What cities in Michigan are hiring for Freelance Site Reliability Engineer jobs?

Cities in Michigan with the most Freelance Site Reliability Engineer job openings:

Manager, Cloud Services and Site Reliability

Barracuda Networks Inc.

Ann Arbor, MI โ€ข On-site

Full-time

Medical, Retirement, PTO

This job post hasย expired today.ย Applications are no longer accepted.


Key responsibilities

  • Lead, coach, and develop a high-performing SRE team, setting clear expectations and fostering a culture of ownership and continuous improvement.

  • Drive reliability engineering practices across critical services, including SLOs, SLIs, monitoring, alerting, capacity planning, and service health reporting.

  • Partner with engineering and platform teams to improve the design, operation, and scalability of cloud-based systems, focusing on reliability, resilience, and maintainability.


Job description

Job ID: 27-0399
Come join our passionate team! Barracuda is a leading cybersecurity company providing complete protection against complex threats. Our platform protects email, data, applications, and networks with innovative solutions, and a managed XDR service, to strengthen cyber resilience. Hundreds of thousands of IT professionals and managed service providers worldwide trust us to protect and support them with solutions that are easy to buy, deploy, and use.

We know a diverse workforce adds to our collective value and strength as an organization. Barracuda Networks is proud to be an Equal Opportunity Employer, committed to equal employment opportunity and equitable compensation regardless of race, gender, religion, sex, sexual orientation, national origin, or disability.

Envision yourself at Barracuda 

We are looking for a Manager, Site Reliability Engineering to lead a team responsible for the reliability, availability, scalability, and operational excellence of high-volume, business-critical SaaS applications. This role combines people leadership with strong technical judgment, helping the team improve reliability practices, reduce operational toil, and support resilient customer-facing services.

In this role, you will manage and develop SRE talent, partner closely with engineering, product, platform, and security teams, and help drive measurable improvements in service health, incident response, automation, and operational readiness. The application portfolio includes products such as Email Security Gateway and Cloud Email Archiving.

What you will be working on
  • Lead, coach, and develop a high-performing SRE team, setting clear expectations, supporting career growth, and fostering a culture of ownership, collaboration, and continuous improvement.

  • Drive reliability engineering practices across critical services, including SLOs, SLIs, monitoring, alerting, capacity planning, and service health reporting.

  • Partner with engineering and platform teams to improve the design, operation, and scalability of cloud-based systems, with a focus on reliability, resilience, and maintainability.

  • Own and improve incident management practices, including major incident coordination, post-incident reviews, follow-up actions, and systemic reliability improvements.

  • Champion automation and tooling that reduce manual effort, improve operational consistency, and help the team scale support for production services.

  • Use operational data, service metrics, and risk indicators to identify reliability gaps, prioritize improvements, and communicate progress to technical and business stakeholders.

  • Support secure and compliant operations by partnering with security and engineering teams to embed appropriate controls, documentation, and operational practices into service delivery.

What you will bring to the role
  • 5+ years of experience in SRE, DevOps, infrastructure, cloud operations, or a related technical operations discipline, including experience leading or managing technical teams.

  • Strong understanding of cloud platforms, distributed systems, production operations, and modern reliability practices.

  • Experience implementing or improving SLOs, SLIs, monitoring, alerting, incident response, and post-incident review practices.

  • Demonstrated ability to hire, mentor, coach, and develop engineers while building a healthy, accountable, and inclusive team culture.

  • Strong communication skills, with the ability to explain technical topics clearly to engineering partners, product stakeholders, and business leaders.

  • Track record of using data, operational insight, and structured problem solving to improve service reliability and team effectiveness.

  • Experience with infrastructure automation, CI/CD practices, disaster recovery, cost optimisation, or multi-cloud operations.

  • Experience influencing operational change across teams, improving documentation practices, or evaluating tools and vendors that support service reliability.

What you’ll get from us 

A team where you can voice your opinion, make an impact, and where you and your experience are valued. Internal mobility – there are opportunities for cross training and the ability to attain your next career step within Barracuda.

  • Equity, in the form of non-qualifying options

  • High-quality health benefits

  • Retirement Plan with employer match

  • Career-growth opportunities

  • Flexible Time Off and Paid Time Off benefits

  • Volunteer opportunities

The anticipated salary range for this role is $151,000 to $200,000. Actual compensation offered will be dependent upon the individual's skills, experience, and qualifications as they directly relate to the requirements of the position, the budget for the position, and applicable employment laws. 

At Barracuda, we believe in fair and equitable compensation practices that reflect both market realities and the unique circumstances of each geographical location. We recognize that cost-of-living disparities, market conditions, and other factors can significantly impact compensation expectations in different regions. The compensation range provided in this job description is for illustrative purposes only and may not reflect the actual compensation offers for the position in your location. Final compensation will be determined based on a variety of factors including the candidates’ qualifications and experience.
#LI-Remote