1

Site Reliability Engineer Intern Jobs in Boston, MA

Cambridge, MA -- Kendall Square HQ (In-Office) The Role As a Senior Site Reliability Engineer at Blitzy's Kendall Square headquarters, you will be a foundational force behind the reliability ...

Site Reliability Engineer

Cambridge, MA ยท On-site

$62.75 - $83.50/hr

The Site Reliability Engineer will be responsible for engineering infrastructure to support high-performance computing needs in a bare metal environment, while also enabling technology evolution and ...

Site Reliability Engineer

Cambridge, MA ยท On-site

$76 - $136/hr

Join our Compute Site Reliability Engineering team Our team is responsible for improving the reliability, performance, and scalability of our Compute products and platforms. We solve complex problems ...

New

Mid Level SRE

Westwood, MA ยท On-site

$63.75 - $84.75/hr

Job Summary : 1872 Consulting is seeking a Site Reliability Engineer to enhance and optimize their infrastructure in Azure, aiming to deliver scalable and reliable software to a global customer base.

Staff Site Reliability Engineer

Cambridge, MA ยท On-site

$160K - $225K/yr

As a Staff SRE, you'll work closely with Platform engineering and Professional Services teams to ensure Manifold's internal and customer-facing infrastructure is secure, scalable, and observable.

Lead Site Reliability Engineer

Boston, MA ยท On-site

$62 - $82.25/hr

The Crown Is Yours As a Lead Site Reliability Engineer, you'll set the reliability standard across our Infrastructure Engineering organization. You'll define how we measure reliability for critical ...

Lead Site Reliability Engineer

Boston, MA ยท On-site

$62 - $82.25/hr

The Crown Is Yours As a Lead Site Reliability Engineer, you'll set the reliability standard across our Infrastructure Engineering organization. You'll define how we measure reliability for critical ...

Staff Site Reliability Engineer

Cambridge, MA ยท Remote

$160K - $225K/yr

As a Staff SRE, you'll work closely with Platform engineering and Professional Services teams to ensure Manifold's internal and customer-facing infrastructure is secure, scalable, and observable.

Site Reliability Engineer II

Cambridge, MA ยท On-site

$62.25 - $82.75/hr

Do you want to build your SRE career on one of the most exciting platforms in cloud computing? Join the Akamai Inference Cloud Team The Akamai Inference Cloud team is part of Akamai's Cloud ...

Showing results 21-40

Site Reliability Engineer Intern information

See Boston, MA salary details

$11

$69

$99

How much do site reliability engineer intern jobs pay per hour?

As of Sep 5, 2026, the average hourly pay for site reliability engineer intern in Boston, MA is $69.25, according to ZipRecruiter salary data. Most workers in this role earn between $59.52 and $79.13 per hour, depending on experience, location, and employer.

What is a site reliability engineer intern?

A Site Reliability Engineer (SRE) Intern supports the reliability, scalability, and performance of software systems by assisting with automation, monitoring, and incident response. They work closely with development and operations teams to improve system resilience and efficiency. Typical tasks include writing scripts, analyzing system logs, and contributing to documentation or tooling improvements. This role helps interns gain hands-on experience in infrastructure management, cloud services, and DevOps practices.

What types of projects or tasks does a site reliability engineer intern typically work on?

As a Site Reliability Engineer Intern, you may work on projects like automating operational processes, building monitoring dashboards, troubleshooting incidents, and helping improve infrastructure reliability. Interns often collaborate closely with full-time engineers on tasks such as writing scripts to automate deployments, responding to on-call alerts, or optimizing system performance. You'll gain practical experience using industry-standard tools and practices while learning about uptime, scalability, and reliability workflows. It's a great opportunity to develop hands-on skills and understand how large-scale systems are maintained in a professional environment.

What are the key skills and qualifications needed to thrive as a site reliability engineer intern, and why are they important?

To thrive as a Site Reliability Engineer Intern, you should have a solid understanding of computer science fundamentals, programming (such as Python or Go), and basic networking concepts, often supported by relevant coursework or prior technical internships. Familiarity with cloud platforms (like AWS or GCP), containerization tools (e.g., Docker, Kubernetes), and version control systems (Git) is highly valuable. Strong problem-solving abilities, effective communication, and a collaborative mindset help interns stand out in team settings. These skills are crucial for learning quickly, contributing to the team's reliability initiatives, and ensuring highly available and scalable system operations.

What are the most commonly searched types of Site Reliability Engineer jobs in Boston, MA?

The most popular types of Site Reliability Engineer jobs in Boston, MA are:

What are popular job titles related to Site Reliability Engineer Intern jobs in Boston, MA?

For Site Reliability Engineer Intern jobs in Boston, MA, the most frequently searched job titles are:

What job categories do people searching Site Reliability Engineer Intern jobs in Boston, MA look for?

The top searched job categories for Site Reliability Engineer Intern jobs in Boston, MA are:

What cities near Boston, MA are hiring for Site Reliability Engineer Intern jobs?

Cities near Boston, MA with the most Site Reliability Engineer Intern job openings:

Infographic showing various Site Reliability Engineer Intern job openings in Boston, MA as of August 2026, with employment types broken down into 1% As Needed, 81% Full Time, 16% Part Time, and 2% Contract. Highlights an 93% Physical, 3% Hybrid, and 4% Remote job distribution, with an average salary of $144,038 per year, or $69.2 per hour.

Senior Site Reliability Engineer

Blitzy

Cambridge, MA โ€ข On-site

$120 - $150/hr

Other

Re-posted 2 days ago


Job description

About this position

Blitzy is a Cambridge, MA based AI software development platform on a mission to revolutionize the software development life cycle by autonomously building custom software to unlock the next industrial revolution. We're transforming how enterprises build software, turning enterprise requirements into production-ready code with an agentic software development platform that can autonomously execute 80% of the quantum of software development work. We're backed by multiple tier 1 investors, and have proven success as founders of previous start-ups.

Location: Cambridge, MA โ€” Kendall Square HQ (In-Office)

The Role

As a Senior Site Reliability Engineer at Blitzy's Kendall Square headquarters, you will be a foundational force behind the reliability, scalability, and operational excellence of our AI-powered software development platform. Sitting at the intersection of software engineering and infrastructure, you'll ensure that the systems enabling enterprise customers to autonomously build production-ready software remain performant, resilient, and always available. This is a high-ownership, high-impact role for an engineer who operates with urgency, thinks in systems, and takes pride in building infrastructure that doesn't break.

What Success Looks Like
  • Blitzy's platform maintains industry-leading uptime โ€” incidents are rare, and when they occur, they are resolved quickly with clear root cause analysis and lasting fixes.
  • SLOs and error budgets are defined for every critical service and actively used to drive engineering decisions, not just tracked passively.
  • Observability is a first-class capability โ€” engineers across the company have the dashboards, traces, and alerts they need to understand system behavior without asking SRE.
  • Deployment pipelines are fast, safe, and reliable โ€” releases go out with confidence and rollbacks are automated when something goes wrong.
  • Infrastructure is entirely codified โ€” no manual provisioning, no configuration drift, every environment reproducible from source.
  • Engineering teams are more productive because of your work โ€” platform friction is low, developer tooling is sharp, and SRE is seen as an accelerant, not a gatekeeper.
  • You are a trusted technical leader at HQ, influencing how Blitzy thinks about reliability as we scale our platform and our team.
Areas of Ownership
  • Design, build, and operate highly available, fault-tolerant infrastructure across cloud environments supporting Blitzy's AI platform and enterprise customers.
  • Define and own SLOs, SLAs, and error budgets for critical services; lead blameless postmortems and drive systemic improvements that prevent recurrence.
  • Build and maintain robust CI/CD pipelines, release automation, and deployment infrastructure that empower engineers to ship with speed and safety.
  • Own the full observability stack โ€” logging, metrics, distributed tracing, and alerting (e.g., Prometheus, Grafana, Datadog, OpenTelemetry).
  • Manage Kubernetes clusters and container infrastructure supporting AI agent workloads and production application services.
  • Drive infrastructure-as-code practices using Terraform; ensure all provisioning is automated, auditable, and version-controlled.
  • Partner with engineering teams at HQ to embed reliability and operational best practices early in the development lifecycle.
  • Lead capacity planning, performance benchmarking, and cloud cost optimization as the platform scales.
Required Experience
  • 5โ€“8 years of experience in Site Reliability Engineering, DevOps, or Platform Engineering.
  • Deep expertise in Kubernetes โ€” cluster management, workload deployment, scaling strategies, and troubleshooting in production.
  • Strong proficiency with at least one major cloud platform (AWS preferred); experience designing and operating distributed, high-availability systems.
  • Hands-on Terraform experience for infrastructure-as-code provisioning and management.
  • Proven ability to define and operationalize SLOs, SLAs, and incident response processes.
  • Strong scripting and automation skills in Python, Go, or Bash.
  • Experience designing and maintaining comprehensive observability systems across complex, multi-service environments.
  • Excellent cross-functional communication skills โ€” able to partner with software engineers, product teams, and leadership equally well.
What Makes You Stand Out
  • Experience operating infrastructure for AI or ML workloads, including GPU scheduling or model serving infrastructure.
  • Familiarity with MLOps tooling (MLflow, Kubeflow, or similar) and the operational challenges unique to AI-driven services.
  • Knowledge of service mesh technologies (Istio, Linkerd) and advanced networking patterns.
  • CKA (Certified Kubernetes Administrator) certification or equivalent demonstrated expertise.
  • Prior experience at a high-growth startup where you built reliability foundations from the ground up.
  • A track record of influencing engineering culture โ€” not just fixing infrastructure, but raising the bar for how teams think about reliability.
What Makes This Role Different

Most SRE roles have you defending the status quo. At Blitzy, you're building reliability infrastructure for a platform that is actively rewriting how enterprises create software โ€” there is no playbook, and that's the point. You'll be based at our Kendall Square headquarters, working daily alongside our co-founders and core engineering team, with direct influence over how we architect and operate systems at the frontier of AI. You'll receive meaningful equity, giving you real ownership in a company that is defining a new category. If you want to do the most consequential infrastructure work of your career, this is the role.

Our Culture

Who we are:

Led by two pioneering co-founders we are one of the fastest growing companies in the U.S., creating our own category of enterprise autonomous software development. We automate thousands of hours of software development for our customers, which includes strong representation within the Fortune 500.

How we work:

We move Blitzy Fast: Time is both our company's and our clients' most precious asset. We move quickly and decisively to innovate internally and deliver exceptional software externally.

Championship Mindset: We operate like a professional sports team. We win as a team by holding ourselves and each other to high standards, collaborating in-person, and remaining focused on the mission.

Passion for Invention: We're pushing the frontier of what's possible, requiring constant innovation and iteration.

We Work for the Customer: We focus on delivering outsized value to the customers we work with and expanding those relationships into deep, meaningful partnerships.

We believe in being 'everyday athletes'โ€”taking care of ourselves so we can bring our best minds to work. We promote great sleep, movement, and restorative activities for optimal mental performance. It makes for a happier and more productive team.

Blitzy is an equal opportunity employer committed to building a diverse and inclusive team. We believe different perspectives make us stronger.

#J-18808-Ljbffr