1

Linux Site Reliability Engineer Jobs in Massachusetts

As a Staff SRE, you'll work closely with Platform engineering and Professional Services teams to ensure Manifold's internal and customer-facing infrastructure is secure, scalable, and observable.

Lead Site Reliability Engineer

Boston, MA

$62 - $82.25/hr

The Crown Is Yours As a Lead Site Reliability Engineer, you'll set the reliability standard across our Infrastructure Engineering organization. You'll define how we measure reliability for critical ...

Lead Site Reliability Engineer

Boston, MA · On-site

$62 - $82.25/hr

The Crown Is Yours As a Lead Site Reliability Engineer, you'll set the reliability standard across our Infrastructure Engineering organization. You'll define how we measure reliability for critical ...

Lead Site Reliability Engineer

Boston, MA · On-site

$62 - $82.25/hr

The Crown Is Yours As a Lead Site Reliability Engineer, you'll set the reliability standard across our Infrastructure Engineering organization. You'll define how we measure reliability for critical ...

Staff Site Reliability Engineer- Eng

Lowell, MA · On-site

$56.50 - $75.25/hr

About the Team Staff Site Reliability Engineers (SREs) at UKG are senior individual contributors ... Strong working knowledge of Linux systems, including troubleshooting, performance analysis, and ...

Site Reliability Engineer

Cambridge, MA · On-site

$62.25 - $82.75/hr

Role Linux Systems Engineers at Watershed engineer infrastructure to meet the high-performance computing needs of modern biology. Designing systems that keep pace with the explosive growth in the ...

Our SRE teams solve reliability, security, and usability at scale for our global fleet while maintaining Akamai's mission at the forefront of what we do: make life better for billions of people ...

Site Reliability Engineer

Cambridge, MA · On-site

$62.25 - $82.75/hr

Role Linux Systems Engineers at Watershed engineer infrastructure to meet the high-performance computing needs of modern biology. Designing systems that keep pace with the explosive growth in the ...

Staff Site Reliability Engineer

Newton, MA · On-site

$62.50 - $83/hr

They are seeking a Staff Site Reliability Engineer to design, build, and operate AWS infrastructure for their platform, ensuring it is secure, scalable, and observable. Responsibilities : • Design ...

Sr. Site Reliability Engineer I

Boston, MA

$62 - $82.25/hr

Your Impact As a Senior Site Reliability Engineer within the APX SRE organization, you'll focus on delivering practical, scalable solutions to support the reliability and performance of our mission ...

Showing results 41-60

Linux Site Reliability Engineer information

What is a Linux Site Reliability Engineer?

A Linux Site Reliability Engineer (SRE) is an IT professional responsible for ensuring the reliability, scalability, and performance of systems running on the Linux operating system. They bridge the gap between software development and operations by automating processes, monitoring infrastructure, and managing incidents. Linux SREs focus on system availability, building tools for deployment and monitoring, and improving system robustness through best practices and automation. Their work helps organizations deliver reliable online services and quickly recover from outages or system failures.

What are the key skills and qualifications needed to thrive as a Linux Site Reliability Engineer?

To thrive as a Linux Site Reliability Engineer, you need deep expertise in Linux system administration, scripting (such as Bash or Python), and a solid understanding of networking concepts, usually backed by a computer science degree or equivalent experience. Familiarity with configuration management tools (like Ansible, Puppet, or Chef), containerization (Docker, Kubernetes), and cloud platforms (AWS, GCP, or Azure) is typically required, along with relevant certifications like RHCE or AWS Certified SysOps Administrator. Strong problem-solving skills, effective communication, and the ability to work under pressure are crucial soft skills for this role. These competencies ensure the reliability, scalability, and security of complex infrastructure, minimizing downtime and supporting seamless operations.

What are some common challenges faced by Linux Site Reliability Engineers when scaling infrastructure, and how can they be addressed?

Linux Site Reliability Engineers often encounter challenges related to maintaining system stability and performance as infrastructure scales. Issues such as configuration drift, automation bottlenecks, and monitoring gaps can arise when managing numerous servers or services. Addressing these challenges typically involves implementing robust configuration management tools, investing in automated deployment pipelines, and enhancing observability through comprehensive monitoring and alerting solutions. Collaboration with development and operations teams is essential to ensure that scalability solutions align with business needs and technical requirements.

What is the difference between Linux Site Reliability Engineer vs Linux DevOps Engineer?

AspectLinux Site Reliability EngineerLinux DevOps Engineer
CredentialsLinux certifications, SRE-specific trainingLinux certifications, DevOps tools certifications
Work EnvironmentFocus on system reliability, monitoring, incident responseFocus on automation, CI/CD pipelines, deployment
Employer & IndustryTech companies, cloud providers, large enterprisesStartups, tech firms, software development teams
Search & Comparison IntentUnderstanding reliability roles, incident managementAutomation, deployment, continuous integration

While both roles involve Linux expertise, a Linux Site Reliability Engineer primarily focuses on maintaining system reliability, monitoring, and incident response. In contrast, a Linux DevOps Engineer emphasizes automation, continuous integration, and deployment processes. Both roles require Linux skills and often overlap, but their core responsibilities differ based on organizational needs.

What are popular job titles related to Linux Site Reliability Engineer jobs in Massachusetts?

For Linux Site Reliability Engineer jobs in Massachusetts, the most frequently searched job titles are:

What job categories do people searching Linux Site Reliability Engineer jobs in Massachusetts look for?

The top searched job categories for Linux Site Reliability Engineer jobs in Massachusetts are:

What cities in Massachusetts are hiring for Linux Site Reliability Engineer jobs?

Cities in Massachusetts with the most Linux Site Reliability Engineer job openings:

Staff Site Reliability Engineer

Cambridge, MA • Remote

$160K - $225K/yr

Full-time

Medical, Dental, Vision, Retirement

Re-posted 4 days ago


Job description

About Manifold

Manifold is the AI platform for life sciences, accelerating life-changing medicines to patients. Our products speed up workflows in areas from target identification and clinical development to market access and precision medicine in the clinic, while maintaining the governance life sciences requires. Global companies and premier research institutions use Manifold to operate faster and more effectively. Backed by leading investors including Reach Capital, TQ Ventures, Calibrate Ventures, SilverArc Capital, and Industry Ventures, Manifold serves tens of thousands of users across hundreds of organizations globally.

About the Role

Manifold is looking for a Staff Site Reliability Engineer (SRE) to work at the intersection of AI, data infrastructure, and life sciences. In this high-impact role, you will help design, build, and operate the multi-account AWS infrastructure that acts as the foundation for Manifold's platform.

As a Staff SRE, you'll work closely with Platform engineering and Professional Services teams to ensure Manifold's internal and customer-facing infrastructure is secure, scalable, and observable. SREs are expected to be fluent across a wide-ranging tech stack, and comfortable working in a high-pressure, multi-threaded environment. They are also expected to balance a bias towards automation with more pragmatic approaches and have an intuitive understanding of the tradeoffs involved. The infrastructure you deploy and manage will play a key role in Manifold's success and help our customers bring life-changing medicines to patients faster.

What You'll Do

  • Design and maintain infrastructure as code solutions, thinking holistically about topology and component dependencies. You will have full responsibility for everything from Terraform plans to production observability.
  • Automate customer infrastructure deployments, including multi-account provisioning, database setup, workflow orchestration, and application bootstrapping.
  • Manage CI/CD pipelines, including build reliability, test stability, and deployment automation for Manifold services.
  • Troubleshoot complex production issues across infrastructure, data, and application layers. Leverage LLM to the fullest extend to minimize toil and and manage date to date operations.
  • Networking and security / compliance work required for highly regulated environments.

What You'll Bring

  • 7+ years in infrastructure, DevOps, SRE, or platform engineering roles with increasing scope and autonomy. You are a leader / doer who can establish operational standards and drive technical direction while also staying hands-on.
  • Deep, hands-on cloud (AWS, GCP, or Azure) experience, hands-on application development experience, and comfortable in troubleshooting application issues.
  • Significant infrastructure-as-code (Terraform) experience. Strong CI/CD (Github Action) experience.
  • Familiarity with identity systems (Okta, Auth0), containerized deployments (Docker, ECS, Packer) and networking tooling (Tailscale, WireGuard). Working knowledge of data platform services, such as Snowflake, Airflow, dbt, and PostgreSQL.
  • Comfort managing complex, multi-account environments where customer isolation, security boundaries, and regulatory requirements add real constraints.
  • You move fast, make sound calls with the information available, and reset quickly when things don't go as planned. And you have strong bias towards pragmatic, incremental process automation.
  • You possess a track record of improving developer experience and reducing CI/CD friction.
  • The ability to effectively and positively collaborate with platform engineer, professional services, and customer IT groups.
  • You've built AI into how you work. You're curious about new tools, resourceful in applying them, and have concrete examples of how AI changed your output.
  • You've done your homework on what it means to accelerate life sciences research and you can articulate why that mission matters to you.
Why Manifold
  • We offer the unique opportunity to work at the frontier of AI in life sciences, helping the world's leading research and clinical institutions adopt transformative technology.
  • This is a high-autonomy, high-impact role where your work directly leads to positive customer outcomes that accelerate breakthrough science.
  • You will gain exposure to a wide range of technical and business challenges across oncology, genomics, clinical data, and AI agent development.
  • You will work with a collaborative team of engineers, scientists, and operators who are building something genuinely new and important.
  • This is an opportunity to grow fast in a company that takes talent seriously. We offer strong compensation packages and excellent benefits.

Salary Range

The base salary range for this position is $160,000-$225,000 annually.

  • Please note: Final offer amounts are determined by multiple factors, including prior experience and expertise, and may vary from the amount above. This range does not represent additional compensation benefits (such as equity, 401K match or medical, dental or vision insurance).