2

Site Reliability Engineer Remote Jobs in Baltimore, MD

Senior Site Reliability Engineer

Millersville, MD · On-site +1

$55.50 - $73.50/hr

The Senior Site Reliability Engineer will apply software engineering principles to operations by automating infrastructure and workflows, defining and measuring reliability targets, strengthening ...

Senior Site Reliability Engineer

Baltimore, MD · On-site +1

$160K - $240K/yr

We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the ... Location - We are flexible on remote working from home, if you are located in the USA and reside in ...

We are currently looking for a Senior Site Reliability Engineer to join our SRE team in the ... Location - We are flexible on remote working from home, if you are located in the USA and reside in ...

Site Reliability Engineer

Columbia, MD · On-site +1

$55.50 - $73.75/hr

Hybrid Columbia MD 3 times per week OR Remote (as applicable to role) Work Authorization ... Job Overview Cogent People Inc. is seeking a Site Reliability to support system reliability ...

DevOps Engineer

Baltimore, MD · Remote

$100K - $120K/yr

This is a remote position. Residency Requirement: Candidates MUST have lived in the U.S. for at ... Collaborate with engineering and SRE stakeholders to reduce toil, improve incident readiness, and ...

You will work closely with the Monitoring & Incident Management Manager, Program Manager, Technical Directors, DevSecOps & SRE teams, and VA stakeholders to identify, escalate, communicate, and help ...

Junior Cloud Engineer

Baltimore, MD · Remote

$50K - $65K/yr

This is a remote position. Residency Requirement: Candidates MUST have lived in the U.S. for at ... Additional certifications in DevOps or SRE practices (e.g., Docker Certified Associate, Certified ...

Senior Data Engineer (Remote)

Baltimore, MD · Remote

$105K - $143K/yr

The Senior Data Engineer is responsible for orchestrating, deploying, maintaining and scaling cloud ... Applies and implements best practices for data auditing, scalability, reliability and application ...

Senior Data Engineer (Remote)

Baltimore, MD · Remote

$105K - $143K/yr

The Senior Data Engineer is responsible for orchestrating, deploying, maintaining and scaling cloud ... Applies and implements best practices for data auditing, scalability, reliability and application ...

next page

Showing results 1-20

Site Reliability Engineer Remote information

See Baltimore, MD salary details

$10

$63

$91

How much do site reliability engineer remote jobs pay per hour?

As of Jul 31, 2026, the average hourly pay for site reliability engineer remote in Baltimore, MD is $63.34, according to ZipRecruiter salary data. Most workers in this role earn between $54.47 and $72.36 per hour, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive in the Site Reliability Engineer Remote position, and why are they important?

To thrive as a Site Reliability Engineer Remote, you need expertise in systems administration, cloud infrastructure, automation, coding (often in Python or Go), and a solid grasp of networking fundamentals, usually demonstrated with a degree in computer science or equivalent experience. Familiarity with tools such as Docker, Kubernetes, AWS/GCP/Azure, monitoring platforms like Prometheus, and certifications like AWS Certified SysOps Administrator are highly valued. Excellent problem-solving, communication, and collaboration skills are essential, especially when troubleshooting incidents and passing information across distributed teams. These abilities ensure reliable, scalable services and smooth coordination in a remote work environment.

What is a Site Reliability Engineer Remote job?

A Site Reliability Engineer (SRE) in a remote role is responsible for ensuring the reliability, performance, and scalability of software systems while working from a remote location. They bridge the gap between development and operations by implementing automation, monitoring, and incident response strategies. Remote SREs collaborate with distributed teams to improve infrastructure, troubleshoot issues, and optimize system performance. Strong communication skills, proficiency in cloud technologies, and expertise in software development are essential for success in this role.

What are some common challenges faced by Site Reliability Engineers working remotely, and how are they addressed?

Site Reliability Engineers working remotely may encounter challenges like coordinating across multiple time zones, maintaining clear communication during urgent incidents, and managing complex systems without direct on-site access. These are often addressed by leveraging collaborative tools (like Slack, Zoom, and incident management platforms), implementing well-documented processes, and participating in regular team syncs or on-call rotations. Remote SREs also benefit from automation and observability practices that provide in-depth systems insights without needing physical presence. Many organizations support their success through robust onboarding, continuous training, and establishing clear lines of communication for rapid response scenarios. This blend of technical and teamwork strategies helps remote SREs maintain service reliability and stay connected with their colleagues.

What are the most commonly searched types of Site Reliability Engineer jobs in Baltimore, MD? The most popular types of Site Reliability Engineer jobs in Baltimore, MD are:
What are popular job titles related to Site Reliability Engineer Remote jobs in Baltimore, MD? For Site Reliability Engineer Remote jobs in Baltimore, MD, the most frequently searched job titles are:
What job categories do people searching Site Reliability Engineer Remote jobs in Baltimore, MD look for? The top searched job categories for Site Reliability Engineer Remote jobs in Baltimore, MD are:
What cities near Baltimore, MD are hiring for Site Reliability Engineer Remote jobs? Cities near Baltimore, MD with the most Site Reliability Engineer Remote job openings:
Infographic showing various Site Reliability Engineer Remote job openings in Baltimore, MD as of July 2026, with employment types broken down into 97% Full Time, 1% Part Time, and 2% Contract. Highlights an 87% Physical, 5% Hybrid, and 8% Remote job distribution, with an average salary of $131,740 per year, or $63.3 per hour.

Senior Site Reliability Engineer

i4DM

Millersville, MD • On-site, Remote

$55.50 - $73.50/hr

Full-time

Re-posted 12 days ago


Job description

Description
About Our Team
Our employees thrive in a culture that is fast-paced, collaborative, and ego-free, where innovation and teamwork are encouraged at every level. We provide Federal agencies with immediate access to highly skilled professionals who understand complex mission challenges and deliver efficient, scalable solutions. By continuously investing in talent, technology, and specialized capabilities, we maintain expert teams prepared to support evolving Federal missions through tailored technical solutions and modern service delivery approaches.
We value diverse perspectives and strive to attract talent from all backgrounds. We are seeking professionals who are passionate about technology, mission success, and solving complex operational challenges with creativity and purpose. If you enjoy expanding your technical expertise while supporting impactful Federal initiatives, you will thrive within our organization. Veterans and military spouses are strongly encouraged to apply and bring their valuable experience to our team.
About the Role
We are seeking an experienced and highly motivated Senior Site Reliability Engineer to serve as a key technical contributor supporting the Technical Director in advancing site reliability engineering, cloud operations, automation, and resilient service delivery for VA enterprise healthcare platforms and applications.
In this role, you will partner closely with the Technical Director, Program Manager, Maintenance Technical Director, Monitoring & Incident Management teams, and VA stakeholders to improve availability, performance, scalability, and operational excellence across mission-critical, 24x7 enterprise environments.
The Senior Site Reliability Engineer will apply software engineering principles to operations by automating infrastructure and workflows, defining and measuring reliability targets, strengthening observability, supporting incident response, and continuously improving system resiliency while aligning with Federal security and governance requirements.
RESPONSIBILITIES
Site Reliability Engineering & Service Ownership
  • Partner with the Technical Director to implement and mature Site Reliability Engineering (SRE) practices across platform services and hosted applications.
  • Improve the full service lifecycle from design and deployment through operation and continuous refinement, with a focus on availability, latency, performance, efficiency, and capacity.
  • Define, track, and report service level indicators (SLIs), service level objectives (SLOs), and error budgets to guide engineering decisions and service improvements.

Automation, CI/CD & Infrastructure as Code
  • Build, enhance, and maintain CI/CD pipelines that enable secure, automated, and repeatable application and infrastructure delivery.
  • Develop and support Infrastructure as Code (IaC) and configuration automation using tools such as Terraform and Ansible to improve consistency, speed, and auditability.
  • Integrate automated testing, validation, and security checks into delivery workflows to improve release quality and reduce change-related risk.

Observability, Reliability & Performance Engineering
  • Design and improve monitoring, logging, tracing, alerting, and dashboards to strengthen observability and accelerate issue detection and response.
  • Analyze system behavior and performance trends to improve reliability, scalability, and operational efficiency across distributed and cloud-native environments.
  • Reduce operational toil by automating repetitive tasks, improving runbooks, and engineering sustainable solutions for recurring operational issues.

Cloud Engineering & Modernization
  • Support cloud infrastructure and platform services in AWS and containerized environments such as Kubernetes, ensuring systems are resilient, scalable, and secure.
  • Contribute to platform modernization efforts by improving deployment patterns, environment consistency, and operational readiness for cloud-native services.
  • Assist with capacity planning, reliability reviews, and architectural improvements to support growth, resilience, and mission continuity.

Security & Compliance Integration
  • Implement reliability engineering practices that align with Federal security requirements, including secure configuration, least privilege, vulnerability remediation, and policy-based controls.
  • Partner with cybersecurity and engineering teams to support secure-by-design infrastructure and application delivery practices.
  • Help ensure operational processes and automation align with compliance expectations for Federal and VA environments.

Cross-Functional Collaboration
  • Collaborate with development, platform, operations, monitoring, incident management, and architecture teams to improve service reliability and deployment outcomes.
  • Work closely with the Technical Director and team leads to translate technical direction into actionable engineering improvements and operational standards.
  • Support Agile and SAFe delivery practices by helping teams adopt reliable release processes, operational readiness checks, and continuous improvement measures.

Incident Support & Continuous Improvement
  • Participate in incident response, service restoration, root cause analysis, and post-incident reviews for critical systems and services.
  • Identify recurring issues, reliability gaps, and failure patterns, and drive corrective actions through automation, architectural improvements, and process refinement.
  • Contribute to on-call readiness, operational documentation, and blameless continuous improvement practices that improve resilience and reduce mean time to recovery.

TAG: #LI-I4DM
TAG: INDMJC
Requirements
QUALIFICATIONS
  • Bachelor's degree in Computer Science, Engineering, Information Technology, or a related technical field, or equivalent practical experience.
  • 5+ years of experience in Site Reliability Engineering, DevOps, platform engineering, cloud operations, or related roles supporting enterprise or mission-critical environments.
  • Hands-on experience supporting cloud platforms (AWS preferred), Linux-based environments, and distributed systems at scale.
  • Strong experience with Infrastructure as Code and automation tools such as Terraform, Ansible, or comparable technologies.
  • Experience with containers and orchestration platforms such as Kubernetes, EKS, ECS, or Docker in production environments.
  • Experience building or maintaining CI/CD pipelines and deployment automation in support of secure, reliable software delivery.
  • Strong understanding of monitoring, observability, incident response, root cause analysis, and performance optimization principles.
  • Proficiency with one or more scripting or programming languages such as Python, Go, Bash, or PowerShell.
  • Demonstrated ability to troubleshoot complex systems, automate operational tasks, and collaborate effectively across engineering and operations teams.
  • Candidates must be eligible to obtain and maintain a Public Trust clearance.

PREFERRED QUALIFICATIONS
  • Experience supporting VA, Federal Government, or other regulated environments with strong security and compliance requirements.
  • Experience defining and operationalizing SLIs, SLOs, error budgets, and service health metrics for production systems.
  • Familiarity with observability platforms and tools such as Prometheus, Grafana, CloudWatch, ELK, Splunk, or OpenTelemetry.
  • Experience with FedRAMP, NIST, Zero Trust, or other Federal security frameworks relevant to cloud and platform operations.
  • Experience supporting healthcare platforms, high-availability enterprise services, or large-scale modernization initiatives.
  • Relevant certifications such as AWS Certified DevOps Engineer, AWS Certified Solutions Architect, Certified Kubernetes Administrator (CKA), HashiCorp Terraform Associate, or SRE/DevOps certifications.

I4dm logo

About I4dm

Sourced by ZipRecruiter

Industry

Software development

Company size

11 - 50 Employees

Headquarters location

Millersville, MD, US

Year founded

2002