1

Site Reliability Engineer Manager Jobs in Baltimore, MD

Design, deploy, and manage AWS infrastructure, including EC2, VPCs, networking, security controls ... Perform Site Reliability Engineering (SRE) functions, including automation of operational tasks ...

Design, deploy, and manage AWS infrastructure, including EC2, VPCs, networking, security controls ... Perform Site Reliability Engineering (SRE) functions, including automation of operational tasks ...

Design, deploy, and manage AWS infrastructure, including EC2, VPCs, networking, security controls ... Perform Site Reliability Engineering (SRE) functions, including automation of operational tasks ...

Site Reliability Engineer

Columbia, MD · On-site

$55.50 - $73.75/hr

They are seeking a Site Reliability Engineer to automate the deployment, scaling, and management of containerized applications, collaborate with developers on CI/CD pipelines, and ensure the stable ...

Site Reliability Engineer

Columbia, MD · On-site

$55.50 - $73.75/hr

Job Overview Cogent People Inc. is seeking a Site Reliability to support system reliability ... Understanding of incident management and root cause analysis processes * Familiarity with cloud ...

Site Reliability Engineer

Columbia, MD · On-site +1

$55.50 - $73.75/hr

Job Overview Cogent People Inc. is seeking a Site Reliability to support system reliability ... Understanding of incident management and root cause analysis processes * Familiarity with cloud ...

Site Reliability Engineer

Columbia, MD · On-site

$55.50 - $73.75/hr

Job Overview Cogent People Inc. is seeking a Site Reliability to support system reliability ... Understanding of incident management and root cause analysis processes * Familiarity with cloud ...

Red Hat/OpenShift SRE

Upper Marlboro, MD · On-site

$56.50 - $75/hr

Respond to and lead incident management for container platform and application-layer issues ... Familiarity with SRE practices: monitoring, alerting, incident response, and blameless post-mortems

next page

Showing results 1-20

Site Reliability Engineer Manager information

See Baltimore, MD salary details

$10

$63

$91

How much do site reliability engineer manager jobs pay per hour?

As of Aug 12, 2026, the average hourly pay for site reliability engineer manager in Baltimore, MD is $63.34, according to ZipRecruiter salary data. Most workers in this role earn between $54.47 and $72.36 per hour, depending on experience, location, and employer.

What is a site reliability engineer manager?

A Site Reliability Engineer (SRE) Manager oversees a team of site reliability engineers tasked with maintaining the reliability, scalability, and performance of software systems. Their role combines leadership and technical expertise, focusing on automating operations, managing incidents, and ensuring high availability of services. They work closely with engineering and operations teams to implement best practices in monitoring, incident response, and system design. SRE Managers also mentor their teams, set reliability goals, and help drive a culture of continuous improvement within the organization.

What is the difference between Site Reliability Engineer Manager vs Site Reliability Engineer?

AspectSite Reliability Engineer (SRE)Site Reliability Engineer Manager
ResponsibilitiesFocuses on designing, implementing, and maintaining reliable systems and automationOversees SRE teams, manages projects, and aligns reliability goals with business objectives
Required SkillsStrong coding, system design, and troubleshooting skillsLeadership, team management, strategic planning
CertificationsGoogle Cloud, AWS certifications, Linux, scriptingSame as SRE, plus management certifications (e.g., PMP) often preferred
Work EnvironmentTechnical, hands-on with systems and automationManagerial, coordinating teams and projects

The main difference is that a Site Reliability Engineer focuses on technical system reliability, while a Site Reliability Engineer Manager oversees teams and strategic initiatives to ensure reliability goals are met across projects.

How does a site reliability engineer manager typically balance technical leadership with team management responsibilities?

A Site Reliability Engineer Manager often splits their time between overseeing technical projects, such as system reliability improvements and incident response strategies, and managing the growth and well-being of their engineering team. This includes mentoring SREs, facilitating communication between teams, setting priorities, and ensuring that operational goals align with business objectives. Balancing these responsibilities requires strong organizational skills and a proactive approach to both technical challenges and people management. Successful managers regularly engage in hands-on problem-solving while also fostering a collaborative team environment.

What are the key skills and qualifications needed to thrive as a site reliability engineer manager?

To thrive as a Site Reliability Engineer Manager, you need expertise in systems engineering, incident management, and a strong background in software development or computer science, often supported by a bachelor’s degree or equivalent experience. Familiarity with cloud platforms (like AWS, GCP, or Azure), infrastructure as code tools (such as Terraform), monitoring systems (like Prometheus), and certifications in cloud or DevOps practices are highly valued. Strong leadership, effective communication, and problem-solving abilities help you guide teams and foster collaboration across departments. These skills and qualities ensure the stability, scalability, and reliability of critical systems while enabling teams to respond effectively to complex technical challenges.
What are the most commonly searched types of Site Reliability Engineer jobs in Baltimore, MD? The most popular types of Site Reliability Engineer jobs in Baltimore, MD are:
What cities near Baltimore, MD are hiring for Site Reliability Engineer Manager jobs? Cities near Baltimore, MD with the most Site Reliability Engineer Manager job openings:

Director of Site Reliability Engineering (SRE)

T Rowe Price

Baltimore, MD • Hybrid

$56.75 - $75.25/hr

Full-time

Re-posted 10 days ago


T. Rowe Price rating

9.1

Company rating: 9.1 out of 10

Based on 21 frontline employees who took The Breakroom Quiz


Job description

Role Summary

The Director of Site Reliability Engineering (SRE) is a strategic leadership role responsible for ensuring the reliability, scalability, and performance of T. Rowe Price's global technology infrastructure & services. This position involves leading a team of skilled engineers to ensure the continuous availability and optimal performance of systems and services, while driving innovation and efficiency in operations.

This position plays a critical role in ensuring overall resilience and efficiency of T. Rowe Price's technology infrastructure, contributing to the organization's success and growth. This position offers an exciting opportunity to lead dynamic teams and drive impactful change in a fast-paced environment.

Responsibilities

Leadership and Strategy:

  • Develop and execute overall strategy for site reliability engineering and observability for the enterprise.
  • Lead SRE and observability teams to ensure alignment with organizational goals and objectives.
  • Establish proactive approach to service reliability that drives reduction in downtime for critical platforms and services.
  • Lead, mentor, and grow a high-performing team of site reliability engineers, fostering a culture of collaboration, innovation, and continuous improvement.
  • Collaborate with cross-functional teams, including software development, IT operations, architecture, security, and product management, to ensure seamless integration and delivery of services.

Reliability and Performance:

  • Oversee the design, implementation, and maintenance of systems and processes that ensure high availability, reliability, and performance of services.
  • Establish and monitor key performance indicators (KPIs) and service level objectives (SLOs) to measure and improve system reliability.
  • Proactively identify and mitigate risks to system reliability, including capacity planning, incident management, and disaster recovery.
  • Lead cross-functional efforts spanning development, engineering, architecture, and operations to identify root cause of instability and drive short/long term improvements.

Automation and Efficiency:

  • Drive the adoption of automation tools and practices to enhance operational efficiency and reduce manual intervention.
  • Implement and refine processes for continuous integration and continuous deployment (CI/CD) to accelerate delivery, improve developer experience and minimize downtime.
  • Promote the use of infrastructure as code (IaC) and other modern practices to streamline operations and improve scalability.

Innovation and Improvement:

  • Stay abreast of industry trends and emerging technologies to identify opportunities for innovation and improvement.
  • Lead initiatives to optimize system architecture and infrastructure, ensuring scalability and adaptability to future needs.
  • Foster a culture of experimentation and learning, encouraging the team to explore new solutions and approaches.

Communication and Collaboration:

  • Serve as a key point of contact for stakeholders, providing regular updates on system reliability and performance.
  • Facilitate effective communication and collaboration between the SRE team and other departments to ensure alignment and shared understanding.
  • Advocate for best practices in site reliability engineering across the organization.

Qualifications

Required:

  • Bachelor's or Master's degree (or the equivalent combination of education and relevant experience and 12+ years in site reliability engineering, cloud engineering, software development, and/or IT operations, with 5+ years of demonstrated technical leadership and team management.
  • Established success implementing and scaling DevOps practices, tools, and frameworks. (i.e. IaC, CI/CD, Agile)
  • Expertise of cloud computing, distributed systems, and modern infrastructure technologies.
  • Proven experience deploying / supporting cloud platforms at scale (AWS, Azure, GCP).
  • Extensive experience deploying observability platforms (i.e Dynatrace, Grafana, Splunk, Dynatrace, OpenTelemetry).
  • Excellent problem-solving skills and the ability to make data-driven decisions.
  • Exceptional communication and interpersonal skills, with the ability to influence and collaborate effectively across teams.

Preferred:

  • Demonstrated expertise in financial services, specifically investment management
  • Certifications: SRECP, CRP

FINRA Requirements

FINRA licenses are not required and will not be supported for this role.

Work Flexibility

This role is eligible for hybrid work, with up to three days per week from home.


What T. Rowe Price employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom