1

Linux Site Reliability Engineer Jobs in Massachusetts

Senior II Site Reliability Engineer

Cambridge, MA ยท On-site

$216K/yr

  • Medical

  • Life

  • PTO

Join our Cloud Networking SRE Team! The Cloud Networking SRE team is part of Akamai ... Demonstrate deep expertise with Linux networking fundamentals and diagnosing at the packet level ...

Site Reliability Engineer

Waltham, MA ยท Hybrid

$40 - $45.78/hr

Site Reliability Engineer 1 Job Details * Site Reliability Engineer 1 (Contract) * Location: Waltham, MA 02451 (Hybrid) * Duration: 10/22/2025 to 4/03/2026 * Team: Campaign Core RD US Key ...

Sr. Site Reliability Engineer

Waltham, MA ยท On-site

$130 - $140/hr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

Site Reliability Engineer****Location(s)**: Waltham, MA | Hybrid**About the Role**Sr Site ... Strong work experience in Unix/Linux* Strong knowledge of Java Web-based enterprise applications ...

Sr. Manager, SRE

Webster, MA ยท On-site

$140K - $185K/yr

  • Medical

  • Retirement

  • PTO

A Site Reliability Engineering (SRE) Senior Manager is responsible for the vision, strategy, and practice of SRE across the organization. Owns reliability, availability, performance, and scalability ...

Sr. Manager, SRE

Webster, MA ยท On-site

$140K - $185K/yr

  • Medical

  • Retirement

  • PTO

A Site Reliability Engineering (SRE) Senior Manager is responsible for the vision, strategy, and practice of SRE across the organization. Owns reliability, availability, performance, and scalability ...

New

Sr. Site Reliability Engineer

Waltham, MA ยท On-site

$61.50 - $81.75/hr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

Site Reliability Engineer Location(s) : Waltham, MA | Hybrid About the Role Sr Site Reliability ... Strong work experience in Unix/Linux * Strong knowledge of Java Web-based enterprise applications ...

Sr. Site Reliability Engineer

Waltham, MA ยท On-site

$61.50 - $81.75/hr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

Site Reliability Engineer Location(s) : Waltham, MA | Hybrid About the Role Sr Site Reliability ... Strong work experience in Unix/Linux * Strong knowledge of Java Web-based enterprise applications ...

Sr. Site Reliability Engineer

Waltham, MA ยท Hybrid

$61.50 - $81.75/hr

  • Medical

  • Dental

  • Vision

  • Retirement

  • PTO

Site Reliability Engineer Location(s) : Waltham, MA | Hybrid About the Role Sr Site Reliability ... Strong work experience in Unix/Linux * Strong knowledge of Java Web-based enterprise applications ...

Site Reliability Engineer II

Cambridge, MA ยท On-site

$62.25 - $82.75/hr

  • Medical

  • Life

Join our highly skilled Site Reliability Team The Platform & Reliability Engineering team is ... Show fluency working in a UNIX/Linux computing environment * Have familiarity with infrastructure ...

Staff Site Reliability Engineer

Cambridge, MA ยท On-site

$160K - $225K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

As a Staff SRE, you'll work closely with Platform engineering and Professional Services teams to ensure Manifold's internal and customer-facing infrastructure is secure, scalable, and observable.

Staff Site Reliability Engineer

Cambridge, MA ยท Remote

$160K - $225K/yr

  • Medical

  • Dental

  • Vision

  • Retirement

As a Staff SRE, you'll work closely with Platform engineering and Professional Services teams to ensure Manifold's internal and customer-facing infrastructure is secure, scalable, and observable.

Showing results 21-40

Linux Site Reliability Engineer information

What is a Linux Site Reliability Engineer?

A Linux Site Reliability Engineer (SRE) is an IT professional responsible for ensuring the reliability, scalability, and performance of systems running on the Linux operating system. They bridge the gap between software development and operations by automating processes, monitoring infrastructure, and managing incidents. Linux SREs focus on system availability, building tools for deployment and monitoring, and improving system robustness through best practices and automation. Their work helps organizations deliver reliable online services and quickly recover from outages or system failures.

What are the key skills and qualifications needed to thrive as a Linux Site Reliability Engineer?

To thrive as a Linux Site Reliability Engineer, you need deep expertise in Linux system administration, scripting (such as Bash or Python), and a solid understanding of networking concepts, usually backed by a computer science degree or equivalent experience. Familiarity with configuration management tools (like Ansible, Puppet, or Chef), containerization (Docker, Kubernetes), and cloud platforms (AWS, GCP, or Azure) is typically required, along with relevant certifications like RHCE or AWS Certified SysOps Administrator. Strong problem-solving skills, effective communication, and the ability to work under pressure are crucial soft skills for this role. These competencies ensure the reliability, scalability, and security of complex infrastructure, minimizing downtime and supporting seamless operations.

What are some common challenges faced by Linux Site Reliability Engineers when scaling infrastructure, and how can they be addressed?

Linux Site Reliability Engineers often encounter challenges related to maintaining system stability and performance as infrastructure scales. Issues such as configuration drift, automation bottlenecks, and monitoring gaps can arise when managing numerous servers or services. Addressing these challenges typically involves implementing robust configuration management tools, investing in automated deployment pipelines, and enhancing observability through comprehensive monitoring and alerting solutions. Collaboration with development and operations teams is essential to ensure that scalability solutions align with business needs and technical requirements.

What is the difference between Linux Site Reliability Engineer vs Linux DevOps Engineer?

AspectLinux Site Reliability EngineerLinux DevOps Engineer
CredentialsLinux certifications, SRE-specific trainingLinux certifications, DevOps tools certifications
Work EnvironmentFocus on system reliability, monitoring, incident responseFocus on automation, CI/CD pipelines, deployment
Employer & IndustryTech companies, cloud providers, large enterprisesStartups, tech firms, software development teams
Search & Comparison IntentUnderstanding reliability roles, incident managementAutomation, deployment, continuous integration

While both roles involve Linux expertise, a Linux Site Reliability Engineer primarily focuses on maintaining system reliability, monitoring, and incident response. In contrast, a Linux DevOps Engineer emphasizes automation, continuous integration, and deployment processes. Both roles require Linux skills and often overlap, but their core responsibilities differ based on organizational needs.

What are popular job titles related to Linux Site Reliability Engineer jobs in Massachusetts?

For Linux Site Reliability Engineer jobs in Massachusetts, the most frequently searched job titles are:

What job categories do people searching Linux Site Reliability Engineer jobs in Massachusetts look for?

The top searched job categories for Linux Site Reliability Engineer jobs in Massachusetts are:

What cities in Massachusetts are hiring for Linux Site Reliability Engineer jobs?

Cities in Massachusetts with the most Linux Site Reliability Engineer job openings:

Senior Application Support Engineer / Site Reliability Engineer (SRE)

DTCC

Boston, MA โ€ข On-site

$62 - $82.50/hr

Full-time

Medical, Life, Retirement, PTO

Re-posted 10 days ago


Job description


Are you ready to make an impact at DTCC?
Do you want to work on innovative projects, collaborate with a dynamic and supportive team, and receive investment in your professional development? At DTCC, we are at the forefront of innovation in the financial markets. We are committed to helping our employees grow and succeed. We believe that you have the skills and drive to make a real impact. We foster a thriving internal community and are committed to creating a workplace that looks like the world that we serve.
The Information Technology group delivers secure, reliable technology solutions that enable DTCC to be the trusted infrastructure of the global capital markets. The team delivers high-quality information through activities that include development of essential, building infrastructure capabilities to meet client needs and implementing data standards and governance.
Pay and Benefits:
  • Competitive compensation, including base pay and annual incentive
  • Comprehensive health and life insurance and well-being benefits, based on location
  • Pension / Retirement benefits
  • Paid Time Off and Personal/Family Care, and other leaves of absence when needed to support your physical, financial, and emotional well-being.
  • DTCC offers a flexible/hybrid model of 3 days onsite and 2 days remote (onsite Tuesdays, Wednesdays and a third day unique to each team or employee).

The Impact You Will Have in This Role
As a Senior Application Support Engineer, you will help power DTCC's global financial markets infrastructure by ensuring the reliability, availability, and performance of Institutional Trade Processing (ITP) platforms that support cross-border equity and debt trade processing and settlement.
Leveraging Site Reliability Engineering (SRE) principles, you will support a portfolio of 40+ mission-critical applications across a modern ecosystem of AWS, OpenShift Container Platform (OCP), Kafka, IBM MQ, and distributed systems. You will play a key role in driving operational excellence, improving resiliency, reducing operational risk, and advancing automation across critical trade processing platforms.
Working closely with globally distributed teams across Application Development, Infrastructure, Cloud, Network, and Operations, you will help deliver stable, scalable, and highly available services that support DTCC's mission-critical business functions.
Your Primary Responsibilities:
Production Reliability & Incident Management
  • Ensure the availability, stability, and performance of mission-critical applications.
  • Lead incident response, troubleshooting, and root cause analysis (RCA) activities.
  • Drive preventative solutions that improve resiliency and reduce recurring issues.

Operational Excellence & Resiliency
  • Support application recovery, failover, disaster recovery (DR), and business continuity activities.
  • Maintain operational readiness across a portfolio of enterprise applications.
  • Support critical processing schedules and service-level commitments.

Change, Release & Automation
  • Support application deployments, releases, and vendor upgrades.
  • Drive automation initiatives that improve efficiency and reduce operational risk.
  • Enhance monitoring, alerting, observability, and operational tooling.

Collaboration & Governance
  • Partner with Development, Infrastructure, Cloud, Database, Network, Security, and Product teams to ensure platform reliability.
  • Support audit, risk, compliance, and operational control requirements.
  • Build strong partnerships across globally distributed teams to drive continuous improvement.

Qualifications
  • Bachelor's degree preferred or equivalent practical experience
  • 6-8 years of experience supporting enterprise applications in complex production environments

Talents Needed for Success:
Core Technologies
  • Java/J2EE
  • Oracle, SQL, DB2
  • IBM MQ, Kafka
  • Splunk, Grafana, AutoSys
  • ServiceNow and ITIL
  • Linux/Unix
  • AWS

Required Experience
  • Strong understanding of Site Reliability Engineering (SRE) principles, including reliability, observability, automation, and incident prevention.
  • Experience supporting enterprise-scale, distributed applications in production environments.
  • Experience with cloud technologies and modern infrastructure platforms.
  • Understanding of messaging systems, middleware, and integration technologies (IBM MQ, Kafka, IIB, or similar).
  • Exposure to container platforms such as OpenShift, Kubernetes, or Docker.
  • Strong troubleshooting, analytical, and problem-solving skills.

Preferred skills
  • Kafka and event-driven architecture experience.
  • Mainframe exposure (JCL, batch processing, job monitoring).
  • Python or other scripting languages.
  • Financial Services, Capital Markets, or Securities Processing experience.
  • AI-driven automation or operational efficiency initiatives.
  • Enterprise integration and distributed processing platforms.

The salary range is indicative for roles at the same level within DTCC across all US locations. Actual salary is determined based on the role, location, individual experience, skills, and other considerations. We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, sex, gender, gender expression, sexual orientation, age, marital status, veteran status, or disability status. We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.
About Us
With over 50 years of experience, DTCC is the premier post-trade market infrastructure for the global financial services industry. From 20 locations around the world, DTCC, through its subsidiaries, automates, centralizes, and standardizes the processing of financial transactions, mitigating risk, increasing transparency, enhancing performance and driving efficiency for thousands of broker/dealers, custodian banks and asset managers. Industry owned and governed, the firm innovates purposefully, simplifying the complexities of clearing, settlement, asset servicing, transaction processing, trade reporting and data services across asset classes, bringing enhanced resilience and soundness to existing financial markets while advancing the digital asset ecosystem. In 2024, DTCC's subsidiaries processed securities transactions valued at U.S. $3.7 quadrillion and its depository subsidiary provided custody and asset servicing for securities issues from over 150 countries and territories valued at U.S. $99 trillion. DTCC's Global Trade Repository service, through locally registered, licensed, or approved trade repositories, processes more than 25 billion messages annually. To learn more, please visit us at www.dtcc.com or connect with us on LinkedIn, X, YouTube, Facebook and Instagram.
DTCC proudly supports Flexible Work Arrangements favoring openness and gives people freedom to do their jobs well, by encouraging diverse opinions and emphasizing teamwork. When you join our team, you'll have an opportunity to make meaningful contributions at a company that is recognized as a thought leader in both the financial services and technology industries. A DTCC career is more than a good way to earn a living. It's the chance to make a difference at a company that's truly one of a kind.
Learn more about Clearance and Settlement by clicking here.
About the Team
Serves as a dedicated technology resource for advancing DTCC's business opportunities and providing industry thought leadership for leveraging new technology. The goal of this new department is to partner internally with IT, our business and regulatory divisions and externally with clients, regulators, and fintech vendors, to help build new platforms and business models to advance DTCC's mission to support the financial markets.