2

Site Reliability Engineer Remote Jobs in Boston, MA

As a Staff SRE, you'll work closely with Platform engineering and Professional Services teams to ensure Manifold's internal and customer-facing infrastructure is secure, scalable, and observable.

Senior DevSecOps Engineer

Boston, MA · Remote

$124K - $170K/yr

Senior DevSecOps Engineer (Remote-based role that requires US-citizenship) About us Hyperproof is ... in SRE, DevSecOps or Platform engineering roles, with a focus on managing Azure-based ...

next page

Showing results 1-20

Site Reliability Engineer Remote information

See Boston, MA salary details

$11

$69

$99

How much do site reliability engineer remote jobs pay per hour?

As of Jul 26, 2026, the average hourly pay for site reliability engineer remote in Boston, MA is $69.25, according to ZipRecruiter salary data. Most workers in this role earn between $59.52 and $79.13 per hour, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive in the Site Reliability Engineer Remote position, and why are they important?

To thrive as a Site Reliability Engineer Remote, you need expertise in systems administration, cloud infrastructure, automation, coding (often in Python or Go), and a solid grasp of networking fundamentals, usually demonstrated with a degree in computer science or equivalent experience. Familiarity with tools such as Docker, Kubernetes, AWS/GCP/Azure, monitoring platforms like Prometheus, and certifications like AWS Certified SysOps Administrator are highly valued. Excellent problem-solving, communication, and collaboration skills are essential, especially when troubleshooting incidents and passing information across distributed teams. These abilities ensure reliable, scalable services and smooth coordination in a remote work environment.

What is a Site Reliability Engineer Remote job?

A Site Reliability Engineer (SRE) in a remote role is responsible for ensuring the reliability, performance, and scalability of software systems while working from a remote location. They bridge the gap between development and operations by implementing automation, monitoring, and incident response strategies. Remote SREs collaborate with distributed teams to improve infrastructure, troubleshoot issues, and optimize system performance. Strong communication skills, proficiency in cloud technologies, and expertise in software development are essential for success in this role.

What are some common challenges faced by Site Reliability Engineers working remotely, and how are they addressed?

Site Reliability Engineers working remotely may encounter challenges like coordinating across multiple time zones, maintaining clear communication during urgent incidents, and managing complex systems without direct on-site access. These are often addressed by leveraging collaborative tools (like Slack, Zoom, and incident management platforms), implementing well-documented processes, and participating in regular team syncs or on-call rotations. Remote SREs also benefit from automation and observability practices that provide in-depth systems insights without needing physical presence. Many organizations support their success through robust onboarding, continuous training, and establishing clear lines of communication for rapid response scenarios. This blend of technical and teamwork strategies helps remote SREs maintain service reliability and stay connected with their colleagues.

What are the most commonly searched types of Site Reliability Engineer jobs in Boston, MA? The most popular types of Site Reliability Engineer jobs in Boston, MA are:
What are popular job titles related to Site Reliability Engineer Remote jobs in Boston, MA? For Site Reliability Engineer Remote jobs in Boston, MA, the most frequently searched job titles are:
What job categories do people searching Site Reliability Engineer Remote jobs in Boston, MA look for? The top searched job categories for Site Reliability Engineer Remote jobs in Boston, MA are:
What cities near Boston, MA are hiring for Site Reliability Engineer Remote jobs? Cities near Boston, MA with the most Site Reliability Engineer Remote job openings:
Infographic showing various Site Reliability Engineer Remote job openings in Boston, MA as of July 2026, with employment types broken down into 100% Full Time. Highlights an 100% Remote job distribution, with an average salary of $144,038 per year, or $69.2 per hour.
Senior Azure Site Reliability Engineer

Senior Azure Site Reliability Engineer

3B Staffing LLC

Boston, MA • Remote

$62 - $82.25/hr

Full-time

This job post has expired today. Applications are no longer accepted.


Job description

looking for an experienced Senior Azure Site Reliability Engineer to ensure the continuous uptime, data integrity, and compliance of healthcare software applications. Operating within an Azure cloud environment, this individual will serve as the primary line of defense for critical production issues, helping minimize disruption to clinical and administrative workflows. The role functions as the primary on-call responder and requires a proactive technical leader who can rapidly resolve critical incidents, automate system recovery processes, and maintain strict HIPAA security standards.
Title: Senior Azure Site Reliability Engineer
Remote / Hybrid Schedule: 100% remote
Must Haves
• 5+ years of experience in Production Support, Site Reliability Engineering (SRE), or a related support engineering function, including 2+ years supporting healthcare environments.
• Hands-on experience managing and supporting production systems within Microsoft Azure.
• Strong expertise in Azure SQL, T-SQL, and healthcare data exchange standards such as HL7 and FHIR.
• Advanced scripting and automation skills using PowerShell, Azure CLI, Bash, or Python.
• Ability to serve as the primary point of contact for high-severity production incidents and critical escalations.
• Experience leading incident response efforts, coordinating bridge calls, and driving rapid resolution within strict SLA requirements.
• Proven ability to perform root cause analysis (RCA) across infrastructure, application, and database layers.
• Commitment to participating in a 24/7 on-call rotation and supporting mission-critical healthcare applications.
• Working knowledge of HIPAA, HITRUST, and healthcare data privacy and security requirements.
Pluses
• Experience supporting Epic EHR environments, including Interconnect, Bridges, Cogito, or related Epic modules.
• Familiarity with Azure OpenAI Services, AI-powered applications, and Large Language Model (LLM) operations.
• Knowledge of monitoring AI application performance, including prompt/response latency and operational health metrics.
• Experience working with vector databases or AI-driven search and retrieval platforms.
• Exposure to enterprise disaster recovery planning, compliance audits, and healthcare regulatory environments.
Day-to-Day
• Act as the primary support lead for high-severity production outages and critical business escalations.
• Coordinate incident response activities and bridge calls to restore service for mission-critical healthcare applications.
• Perform troubleshooting and root cause analysis across Azure infrastructure, databases, and application services.
• Execute recovery and remediation activities to minimize disruption to patient care and business operations.
• Monitor and maintain the health, availability, and performance of Azure App Services, Virtual Machines, and Azure Functions.
• Analyze telemetry and system performance data using Azure Monitor, Log Analytics, and Application Insights.
• Support and optimize Azure SQL databases, enterprise data pipelines, and cloud storage solutions.
• Deploy system patches, hotfixes, and updates in accordance with security and operational standards.
• Audit system logs, user access, and security controls to ensure ongoing compliance and data protection.
• Develop and maintain automation scripts to reduce manual support activities and improve operational efficiency.
• Create and maintain technical documentation, support runbooks, disaster recovery procedures, and compliance artifacts.
• Collaborate with engineering, infrastructure, security, and application teams to drive service reliability and continuous improvement.