1

Site Reliability Engineer Manager Jobs in Boston, MA

Site Reliability Engineer

Waltham, MA · On-site

$61.50 - $81.75/hr

You'll be the second SRE on a small, high-trust team, working directly with our lead, with real ... Familiarity with systemd service management and observability practices on Linux * Familiarity with ...

Site Reliability Engineer

Waltham, MA · On-site

$61.50 - $81.75/hr

You'll be the second SRE on a small, high-trust team, working directly with our lead, with real ... Familiarity with systemd service management and observability practices on Linux * Familiarity with ...

Site Reliability Engineer

Waltham, MA · Hybrid

$40 - $45.78/hr

Site Reliability Engineer 1 Job Details * Site Reliability Engineer 1 (Contract) * Location ... Deploy and manage solutions efficiently on private or public cloud, and ensure they meet ...

Site Reliability Engineer

Waltham, MA

$61.50 - $81.75/hr

You'll be the second SRE on a small, high-trust team, working directly with our lead, with real ... Familiarity with systemd service management and observability practices on Linux * Familiarity with ...

Site Reliability Engineer

Boston, MA · On-site

$155K - $222K/yr

We are one of several SRE teams working together to support a platform that serves more than 500,000 customers and manages over 18 million devices worldwide. The team operates with a high degree of ...

Site Reliability Engineer

Cambridge, MA · On-site

$75K - $136K/yr

Our team designs, develops, and manages applications and infrastructure that support Akamai Cloud's products and services. Our SRE teams solve reliability, security, and usability at scale for our ...

Senior Technology Site Reliability Engineer

Boston, MA · On-site

$62 - $82.25/hr

Implement and manage service-level indicators (SLIs), objectives (SLO's), agreements (SLA's), and ... g. site reliability engineering or related field) * Proficiency in Terraform and programming ...

Site Reliability Engineer

Cambridge, MA · On-site

$62.75 - $83.50/hr

The Site Reliability Engineer will be responsible for engineering infrastructure to support high ... Responsibilities : • Build, manage, and deploy a large fleet of Linux servers • Research new ...

As a Staff SRE, you'll work closely with Platform engineering and Professional Services teams to ... Manage CI/CD pipelines, including build reliability, test stability, and deployment automation for ...

As a Staff SRE, you'll work closely with Platform engineering and Professional Services teams to ... Manage CI/CD pipelines, including build reliability, test stability, and deployment automation for ...

next page

Showing results 1-20

Site Reliability Engineer Manager information

See Boston, MA salary details

$11

$69

$99

How much do site reliability engineer manager jobs pay per hour?

As of Aug 9, 2026, the average hourly pay for site reliability engineer manager in Boston, MA is $69.25, according to ZipRecruiter salary data. Most workers in this role earn between $59.52 and $79.13 per hour, depending on experience, location, and employer.

What is a site reliability engineer manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.
What are the most commonly searched types of Site Reliability Engineer jobs in Boston, MA? The most popular types of Site Reliability Engineer jobs in Boston, MA are:
What cities near Boston, MA are hiring for Site Reliability Engineer Manager jobs? Cities near Boston, MA with the most Site Reliability Engineer Manager job openings:
Infographic showing various Site Reliability Engineer Manager job openings in Boston, MA as of July 2026, with employment types broken down into 86% Full Time, 12% Part Time, and 2% Contract. Highlights an 84% Physical, 2% Hybrid, and 14% Remote job distribution, with an average salary of $144,037 per year, or $69.2 per hour.

Site Reliability Engineer SRE Weekend Coverage

Framework Ventures

Boston, MA • On-site

$120 - $180/hr

Other

Posted 4 days ago


Job description

Site Reliability Engineer (SRE) – Weekend Coverage

About Elwood: We have built a digital asset trading infrastructure for institutional investors. Our seamless end‑to‑end platform connects to global crypto exchanges, custodians, and liquidity providers, via a single API. Built by industry experts with decades of combined experience in investment management and digital technology, Elwood provides market infrastructure at scale, enabling financial institutions, neobanks, and corporations to access digital asset markets quickly and efficiently.

We are seeking a Site Reliability Engineer (SRE) to join our globally distributed engineering team, with a key responsibility for weekend operations and system reliability. You’ll play a critical role in maintaining uptime, resolving incidents, and automating infrastructure for Elwood’s EMS and PMS platforms, which are built on AWS and GCP cloud environments. This highly visible role blends deep technical ownership with cross‑functional collaboration. In addition to core SRE responsibilities, you will support our Technical Account Managers and client‑facing teams in resolving production issues that impact users, ensuring smooth and reliable client experiences. You’ll be part of a team responsible for both the infrastructure backbone and the performance reputation of Elwood’s platform.

Key Responsibilities
  • Ensure the reliability, availability, and performance of production systems, particularly during weekends.
  • Take ownership of monitoring, troubleshooting, and incident response during weekends and off‑hours.
  • Actively participate in on‑call rotations with priority weekend shifts (Saturday–Sunday).
  • Troubleshoot and resolve critical issues in a fast‑paced, high‑availability environment.
  • Automate manual processes and workflows, reducing operational overhead.
  • Work closely with engineering teams to design and deploy scalable, fault‑tolerant infrastructure solutions on AWS or GCP.
  • Improve observability by utilizing monitoring, logging, and alerting systems (e.g., CloudWatch, Datadog).
  • Lead post‑incident reviews, contribute to the continuous improvement of system reliability, and follow up on strategic fixes.
  • Develop and update runbooks, incident response playbooks, and documentation.
  • Work closely with Engineering, Product, and Client teams to proactively identify infrastructure pain points that could affect the user experience.
  • Monitor alert channels, logs, and infrastructure load for the entire stack.
  • Set up automation for alerting.
Required Experience
  • 5+ years of experience in an SRE, DevOps, or infrastructure engineering role.
  • Strong experience with AWS or GCP, including services like EC2, Lambda, S3, RDS, and GKE (for GCP).
  • Experience with automation tools like Terraform.
  • Proficient in at least one scripting language (Python, Bash, Go, etc.).
  • Solid understanding of Linux systems, networking, and cloud‑based architectures.
  • Experience working with container orchestration platforms like Kubernetes.
  • Proficient with CI/CD pipelines, preferably with cloud‑native tools (e.g., GitHub).
  • Ability to troubleshoot complex, distributed systems and provide solutions in high‑pressure environments.
  • Ability to communicate effectively with both technical and non‑technical stakeholders.
Preferred Qualifications
  • Exposure to Execution Management Systems (EMS) / Portfolio Management Systems (PMS).
  • Experience with client‑impact triage, working cross‑functionally with account managers or product teams.
  • Proficiency with Datadog or similar observability platforms.
  • Knowledge of serverless architectures (e.g., AWS Lambda, GCP Cloud Functions).
  • Familiarity with RDBMS and NoSQL databases, such as RDS, CloudSQL, DynamoDB.
  • Prior experience in fintech, trading platforms, or 24/7 financial infrastructure.
  • Strong understanding of API integrations and how infrastructure issues might manifest in client environments.
  • Excellent problem‑solving and communication skills, with the ability to translate technical incidents into clear client updates.
  • Experience working with client‑facing teams.
Working Hours and Availability

The successful candidate must be comfortable working autonomously during weekend shifts (Thursday‑Monday schedule).

Values
  • Passion – As a Software Engineer, you are passionate about building high‑quality, performant services.
  • Respect – As part of a diverse team, you have deep cultural empathy and respect for fellow workers, clients, partners, and Elwood’s regulatory obligations.
  • Teamwork & Communication – You value cross‑team collaboration, effective communication, and efficient processes that deliver best results.
  • Tenacity – You know that delivering stable, performant systems takes hard work and persistence, and you bring both these qualities to the table every day.
  • Unblock – You are willing to share your knowledge with the rest of the team and aren’t afraid to ask questions to improve your work and progress delivery.
  • Trust & Transparency – You are willing to share your knowledge with the rest of the team and aren’t afraid to ask questions to improve your work and progress delivery.
  • Excellence – Excellence is a mindset and you strive to be the best version of yourself and expect the same from others. You will have a testing mindset and ruthless focus on delivering quality solutions and value for our clients.
Equal Opportunities

As an equal opportunity employer, you can read more about our policy here: https://elwood.io/diversity

#J-18808-Ljbffr