1

Linux Site Reliability Engineer Jobs in Pennsylvania

Site Reliability Engineer

Malvern, PA · On-site

$56 - $74.25/hr

#W2 Role Senior Reliability Engineer As a Senior Reliability Engineer, you will play a critical role in solving impactful operational problems. You are curious and take a proactive approach to ...

Site Reliability Engineering

Philadelphia, PA

$57.50 - $76.50/hr

Serve as subject matter expert in an SRE mindset, best practices, and cloud-native principles * Scale systems sustainably through automation to improve reliability and velocity * Assist with all ...

Senior Site Reliability Engineer

Wayne, PA · On-site

$51.75 - $68.75/hr

Actively participate in reliability engineering and resilience communities of practice, contributing to shared learning and enterprise consistency. * Contribute to strategic initiatives that advance ...

Senior Site Reliability Engineer

Wayne, PA · On-site

$51.75 - $68.75/hr

Actively participate in reliability engineering and resilience communities of practice, contributing to shared learning and enterprise consistency. * Contribute to strategic initiatives that advance ...

Minimum of 5 years of experience in automating infrastructure, service delivery, and engineering site reliability, maintaining infrastructure on premise and in cloud environment * Product does not ...

DevOps Engineer

Pittsburgh, PA · On-site

$51.25 - $70.25/hr

Collaborate with development, QA, SRE, and security teams to improve delivery and operational ... Linux * Python * Shell Script Ref: #404-IT Pittsburgh

DevOps Engineer

Pittsburgh, PA · On-site

$51.25 - $70.25/hr

Collaborate with development, QA, SRE, and security teams to improve delivery and operational ... Linux * Python * Shell Script Ref: #404-IT Pittsburgh

Senior DevOps/SRE Engineer

Oaks, PA · On-site

$140K - $170K/yr

BA/BS, in a related technical field; or the equivalent in education and work experience * 8+ years of experience in DevOps, SRE, platform engineering, or similar roles supporting application teams ...

Showing results 41-60

Linux Site Reliability Engineer information

What is a Linux Site Reliability Engineer?

A Linux Site Reliability Engineer (SRE) is an IT professional responsible for ensuring the reliability, scalability, and performance of systems running on the Linux operating system. They bridge the gap between software development and operations by automating processes, monitoring infrastructure, and managing incidents. Linux SREs focus on system availability, building tools for deployment and monitoring, and improving system robustness through best practices and automation. Their work helps organizations deliver reliable online services and quickly recover from outages or system failures.

What are the key skills and qualifications needed to thrive as a Linux Site Reliability Engineer?

To thrive as a Linux Site Reliability Engineer, you need deep expertise in Linux system administration, scripting (such as Bash or Python), and a solid understanding of networking concepts, usually backed by a computer science degree or equivalent experience. Familiarity with configuration management tools (like Ansible, Puppet, or Chef), containerization (Docker, Kubernetes), and cloud platforms (AWS, GCP, or Azure) is typically required, along with relevant certifications like RHCE or AWS Certified SysOps Administrator. Strong problem-solving skills, effective communication, and the ability to work under pressure are crucial soft skills for this role. These competencies ensure the reliability, scalability, and security of complex infrastructure, minimizing downtime and supporting seamless operations.

What are some common challenges faced by Linux Site Reliability Engineers when scaling infrastructure, and how can they be addressed?

Linux Site Reliability Engineers often encounter challenges related to maintaining system stability and performance as infrastructure scales. Issues such as configuration drift, automation bottlenecks, and monitoring gaps can arise when managing numerous servers or services. Addressing these challenges typically involves implementing robust configuration management tools, investing in automated deployment pipelines, and enhancing observability through comprehensive monitoring and alerting solutions. Collaboration with development and operations teams is essential to ensure that scalability solutions align with business needs and technical requirements.

What is the difference between Linux Site Reliability Engineer vs Linux DevOps Engineer?

AspectLinux Site Reliability EngineerLinux DevOps Engineer
CredentialsLinux certifications, SRE-specific trainingLinux certifications, DevOps tools certifications
Work EnvironmentFocus on system reliability, monitoring, incident responseFocus on automation, CI/CD pipelines, deployment
Employer & IndustryTech companies, cloud providers, large enterprisesStartups, tech firms, software development teams
Search & Comparison IntentUnderstanding reliability roles, incident managementAutomation, deployment, continuous integration

While both roles involve Linux expertise, a Linux Site Reliability Engineer primarily focuses on maintaining system reliability, monitoring, and incident response. In contrast, a Linux DevOps Engineer emphasizes automation, continuous integration, and deployment processes. Both roles require Linux skills and often overlap, but their core responsibilities differ based on organizational needs.

What are popular job titles related to Linux Site Reliability Engineer jobs in Pennsylvania?

For Linux Site Reliability Engineer jobs in Pennsylvania, the most frequently searched job titles are:

What job categories do people searching Linux Site Reliability Engineer jobs in Pennsylvania look for?

The top searched job categories for Linux Site Reliability Engineer jobs in Pennsylvania are:

What cities in Pennsylvania are hiring for Linux Site Reliability Engineer jobs?

Cities in Pennsylvania with the most Linux Site Reliability Engineer job openings:

Data Center Site Reliability Engineer

Synopsys, Inc.

Canonsburg, PA • On-site

$90 - $140/hr

Other

Posted 6 days ago


Job description

Data Center Site Reliability Engineer

Synopsys is the leader in engineering solutions from silicon to systems, enabling customers to rapidly innovate AI-powered products. We deliver industry-leading silicon design, IP, simulation and analysis solutions, and design services. We partner closely with our customers across a wide range of industries to maximize their R&D capability and productivity, powering innovation today that ignites the ingenuity of tomorrow.

You Are

You are someone who takes ownership when things are complex, time-sensitive, and highly visible. Youdon’twait for perfect conditions or completeinformation.Youcreate clarity by asking the right questions, confirming assumptions, and moving the work forward in a way that keeps others aligned and confident. You thrive in environments where priorities can shift quickly, and you staysteadyunder pressure by focusing on what matters most: protecting reliability,maintainingtrustand delivering results that hold up over time.

What You'll Be Doing
  • Serve as the primary on-site technical resource supporting the Canonsburg data center and coordinating day-to-day operational needs with global engineering teams.
  • Perform Linux systems administration tasks, including basic troubleshooting, log analysis, remote access support, and service management to keep critical systems running reliably.
  • Install, rack, cable ,relocate, and decommission physical server and networking equipment while maintaining clean, organized rack layouts and labeling standards.
  • Maintain accurate infrastructure documentation and digital asset records in DCIM tools (e.g., Sunbird or similar), ensuring inventory, connectivity, and capacity data stays current.
  • Coordinate and oversee third-party vendors performing maintenance and infrastructure work, confirming scope, access requirements, safety practices, and completion criteria.
  • Monitor data center health indicators such as power, cooling, rack capacity, and environmental conditions, escalating risks and initiating corrective actions as needed.
  • Respond to operational incidents as part of a shared on-call rotation, meeting established response SLAs and driving issues through resolution and follow-up.
The Impact You Will Have
  • Keep mission-critical engineering systems dependable by reducing downtime and restoring service quickly when issues arise.
  • Improve day-to-day operational confidence through accurate asset records and documentation that make capacity, ownership, and change planning clear.
  • Accelerate infrastructure deployments by ensuring on-site execution is timely, consistent, and aligned with global engineering standards.
  • Reduce operational risk by identifying early warning signs in power, cooling, and environmental conditions and driving corrective actions before they become incidents.
  • Strengthen vendor outcomes by ensuring work is properly scoped, safely executed, and fully completed with clear validation and follow-through.
  • Increase cross-team effectiveness by serving as a reliable on-site partner who communicates clearly, escalates appropriately, and closes loops after changes and incidents.
  • Support successful data center integration and modernization by helping standardize processes and stabilizing operations during periods of change.
What You'll Need
  • You have experience administering Linux systems (Red Hat, Rocky Linux, Ubuntu, or similar) and can navigate common operational tasks with confidence.
  • You bring hands-on familiarity working in a physical enterprise data center environment, where safety, precision, and process matter.
  • You have a working understanding of server hardware, rack infrastructure, structured cabling, power distribution, and cooling fundamentals.
  • You bring strong troubleshooting and problem-solving habits, including the ability to stay calm, prioritize effectively, and drive issues to resolution.
  • You have experience coordinating with third-party vendors and service providers and can ensure work is completed to scope and standard.
  • You are able to work independently as the primary on-site technical resource and communicate clearly with remote engineering partners.
  • You bring differentiators such as exposure to DCIM platforms (e.g., Sunbird), network/storage/virtualization environments, and/or scripting and automation with Bash, Python, or PowerShell.
Who You Are
  • You are the kind of person who takes ownership end-to-end, following through until the issue is fully resolved and the next steps are clear to everyone involved.
  • You approach problems methodically, separating symptoms from root causes and validating changes before and after you act.
  • You communicate with precision, tailoring updates to the audience and escalating early when risk, safety, or SLA impact is on the line.
  • You stay organized in fast-moving environments, keeping documentation, labels, and records accurate so others can operate confidently after you.
  • You collaborate smoothly across teams and vendors, setting expectations upfront and ensuring work is completed safely, cleanly, and to standard.
The Team You'll Be Part Of

The Data Center Operations team supports reliable, secure day-to-day execution across Synopsys’ physical infrastructure footprint while partnering closely with global network, storage, security, and infrastructure engineering teams. This role serves as the primary on-site operator for the Canonsburg data center, ensuring operational excellence and enabling modernization and integration efforts.

Rewards and Benefits

We offer a comprehensive range of health, wellness, and financial benefits to cater to your needs. Our total rewards include both monetary and non-monetary offerings.

#J-18808-Ljbffr

Synopsys logo

About Synopsys

Sourced by ZipRecruiter

Synopsys, Inc. (Nasdaq:SNPS) is the Silicon to Software partner for creative companies developing the electronic products and software applications we rely on every single day. As the world's 15th largest software company, Synopsys has a long history of being a global leader in electronic design automation (EDA) and semiconductor IP and is also growing its leadership in software quality and security solutions. Whether you're a system-on-chip (SoC) designer building advanced semiconductors, or a software developer writing applications that require the highest quality and security, Synopsys has the solutions needed to deliver exceptional, secure products for the era of connected everything. The company is headquartered in Mountain View, California, and has approximately 113 offices located throughout North America, South America, Europe, Japan, Asia and India. Since 1986, Synopsys has been at the heart of accelerating electronics innovation with engineers around the world having used Synopsys technology to successfully design and create billions of chips and systems that are found in the electronics that people rely on every single day.

Industry

Computer and computer peripheral equipment and software wholesalers

Company size

10,000+ Employees

Headquarters location

Mountain View, CA, US

Year founded

1986

Social media