2

Site Reliability Engineer Remote Jobs in New Jersey

Champion modern engineering practices including automation, observability, DevOps, Site Reliability ... We embrace a remote-first culture through our Flexible Workplace. Most employees hold Home-Flex ...

DevOps Engineer

Hoboken, NJ · Remote

$57.75 - $79/hr

Experience: 5+ years of experience in DevOps, Site Reliability Engineering (SRE), or a related role ... Benefits We offer a fully remote environment, plus a competitive benefits package including medical ...

DevOps Engineer

Hoboken, NJ · On-site +1

$57.75 - $79/hr

Experience: 5+ years of experience in DevOps, Site Reliability Engineering (SRE), or a related role ... Benefits We offer a fully remote environment, plus a competitive benefits package including medical ...

DevOps Engineer

Hoboken, NJ · Remote

$57.75 - $79/hr

Experience: 5+ years of experience in DevOps, Site Reliability Engineering (SRE), or a related role ... Benefits We offer a fully remote environment, plus a competitive benefits package including medical ...

DevOps Engineer

Cherry Hill, NJ · On-site +1

$52.25 - $71.75/hr

... #SRE #DevOps #programming #kubernetes #security #compliance (ISO 27001, SOC). What are you looking ... Hybrid & remote work. So, what's up? If you liked what you read, don't hesitate.. hit on apply, we ...

next page

Showing results 1-20

Site Reliability Engineer Remote information

See New Jersey salary details

$10

$64

$93

How much do site reliability engineer remote jobs pay per hour?

As of Aug 31, 2026, the average hourly pay for site reliability engineer remote in New Jersey is $64.71, according to ZipRecruiter salary data. Most workers in this role earn between $55.62 and $73.94 per hour, depending on experience, location, and employer.

What is a site reliability engineer remote?

A Site Reliability Engineer (SRE) in a remote role is responsible for ensuring the reliability, performance, and scalability of software systems while working from a remote location. They bridge the gap between development and operations by implementing automation, monitoring, and incident response strategies. Remote SREs collaborate with distributed teams to improve infrastructure, troubleshoot issues, and optimize system performance. Strong communication skills, proficiency in cloud technologies, and expertise in software development are essential for success in this role.

What are the key skills and qualifications needed to thrive as a site reliability engineer remote?

To thrive as a Site Reliability Engineer Remote, you need expertise in systems administration, cloud infrastructure, automation, coding (often in Python or Go), and a solid grasp of networking fundamentals, usually demonstrated with a degree in computer science or equivalent experience. Familiarity with tools such as Docker, Kubernetes, AWS/GCP/Azure, monitoring platforms like Prometheus, and certifications like AWS Certified SysOps Administrator are highly valued. Excellent problem-solving, communication, and collaboration skills are essential, especially when troubleshooting incidents and passing information across distributed teams. These abilities ensure reliable, scalable services and smooth coordination in a remote work environment.

What are some common challenges faced by site reliability engineers working remotely, and how are they addressed?

Site Reliability Engineers working remotely may encounter challenges like coordinating across multiple time zones, maintaining clear communication during urgent incidents, and managing complex systems without direct on-site access. These are often addressed by leveraging collaborative tools (like Slack, Zoom, and incident management platforms), implementing well-documented processes, and participating in regular team syncs or on-call rotations. Remote SREs also benefit from automation and observability practices that provide in-depth systems insights without needing physical presence. Many organizations support their success through robust onboarding, continuous training, and establishing clear lines of communication for rapid response scenarios. This blend of technical and teamwork strategies helps remote SREs maintain service reliability and stay connected with their colleagues.

What are the most commonly searched types of Site Reliability Engineer jobs in New Jersey?

The most popular types of Site Reliability Engineer jobs in New Jersey are:

What job categories do people searching Site Reliability Engineer Remote jobs in New Jersey look for?

The top searched job categories for Site Reliability Engineer Remote jobs in New Jersey are:

What cities in New Jersey are hiring for Site Reliability Engineer Remote jobs?

Cities in New Jersey with the most Site Reliability Engineer Remote job openings:

Infographic showing various Site Reliability Engineer Remote job openings in New Jersey as of August 2026, with employment types broken down into 1% As Needed, 80% Full Time, 15% Part Time, 1% Temporary, 2% Contract, and 1% Nights. Highlights an 94% Physical, 2% Hybrid, and 4% Remote job distribution, with an average salary of $134,603 per year, or $64.7 per hour.

Senior Manager, Site Reliability Engineer - Remote

Basking Ridge, NJ • On-site, Remote


UnitedHealth Group
Insurance Services • 10K+ employees

7.6

Company rating: 7.6 out of 10

Based on 146 frontline employees who took The Breakroom Quiz

192nd of 896 rated healthcare providers

Good employer

Recommended by students

Recommended by parents


$58.75 - $78/hr

Full-time

Retirement

Posted 24 days ago


Job description

Optum Tech is a global leader in health care innovation. Our teams develop cutting-edge solutions that help people live healthier lives and help make the health system work better for everyone. From advanced data analytics and AI to cybersecurity, we use innovative approaches to solve some of health care's most complex challenges. Your contributions here have the potential to change lives. Ready to build the next breakthrough? Join us to start Caring. Connecting. Growing together.
We are seeking an experienced Senior Manager to lead enterprise Site Reliability Engineering (SRE), DevOps, IT Service Management (ITSM), and Operational Excellence initiatives across Optum bank. This leader will be responsible for improving service reliability, operational resiliency, deployment automation, observability, incident management, and production readiness for critical banking platforms.
The ideal candidate combines strong technical expertise with operational leadership experience, driving engineering excellence, automation, reliability, and continuous improvement while ensuring technology services meet business, customer, regulatory, and operational expectations.
This role will also help identify and implement emerging automation and AI-enabled operational capabilities that improve service health, reduce operational toil, and accelerate engineering productivity.
Primary Responsibilities:
  • Leadership & Engineering Management
    • Lead and develop multidisciplinary teams responsible for Site Reliability Engineering, DevOps, Platform Engineering, ITSM, and Operational Excellence
    • Establish and execute enterprise reliability, availability, resiliency, and operational maturity strategies
    • Drive engineering excellence through automation, observability, operational readiness, and continuous improvement practices
    • Partner with Technology, Operations, Security, Infrastructure, Risk, and Business leaders to improve service reliability and customer experience
    • Build and mentor high-performing teams while fostering accountability, innovation, operational ownership, and learning
    • Manage staffing, capacity planning, talent development, succession planning, and organizational growth
    • Establish operational metrics, governance standards, and service review processes to improve service performance and risk management
  • SRE, DevOps & Platform Engineering
    • Lead enterprise SRE practices including SLI/SLO adoption, error-budget management, reliability engineering, and operational maturity assessments
    • Drive DevOps transformation initiatives, emphasizing automation, deployment standardization, CI/CD pipelines, Infrastructure-as-Code, and GitOps practices
    • Establish production readiness standards and operational acceptance criteria for new technology deployments
    • Improve platform resiliency through capacity planning, disaster recovery, fault tolerance, and resilience testing
    • Drive reduction of operational toil through automation and self-healing capabilities
    • Partner with application and infrastructure teams to improve system scalability, availability, and performance
    • Lead initiatives to improve deployment frequency, reduce change failure rates, and accelerate service recovery times
  • ITSM & Operational Excellence
    • Establish and mature Incident, Problem, Change, Release, and Service Request Management processes
    • Lead major incident management programs and executive communications during critical service disruptions
    • Drive root-cause analysis and problem-management practices to eliminate recurring incidents
    • Improve operational scorecards, service health reviews, and reliability reporting for executive stakeholders
    • Ensure compliance with regulatory, audit, risk, and operational governance requirements
    • Partner with Technology and Business leaders to improve service quality and customer outcomes through data-driven operational improvements
    • Champion a culture of operational excellence and continuous service improvement
  • Intelligent Automation & AI-Enabled Operations
    • Identify opportunities to leverage AI and automation to improve operational effectiveness and engineering productivity
    • Lead implementation and evaluation of solutions involving: AIOps, Intelligent alert correlation, Automated incident triage , Root cause analysis assistance, Knowledge management copilots, Agentic operational workflows etc.
    • Partner with enterprise AI teams to evaluate emerging technologies that improve reliability and operational efficiency
    • Drive responsible adoption of AI-enabled engineering and operational practices
    • Support proof-of-concept initiatives that demonstrate measurable reductions in operational effort and incident resolution times
  • Cross-Functional Leadership
    • Collaborate with Engineering, Infrastructure, Security, Architecture, Risk, Compliance, and Operations teams to prioritize reliability and operational improvements
    • Serve as a trusted advisor on reliability engineering, operational excellence, and automation strategies
    • Drive alignment between technology and business stakeholders to improve service quality and operational outcomes
    • Influence technology investment decisions that improve platform stability, resiliency, and operational efficiency

You'll be rewarded and recognized for your performance in an environment that will challenge you and give you clear direction on what it takes to succeed in your role as well as provide development for other roles you may be interested in.
Required Qualifications:
  • Bachelor's degree in Computer Science, Engineering, Information Technology, or related field
  • 10+ years of experience in Software Engineering, Site Reliability Engineering, Platform Engineering, DevOps, Infrastructure Engineering, or Technology Operations
  • 5+ years of experience leading engineering or operational teams
  • Proven experience supporting large-scale, business-critical production environments
  • Experience with:
    • SRE principles and practices
    • DevOps and CI/CD
    • ITSM processes
    • Cloud platforms (Azure, AWS)
    • On-prem environments
    • Infrastructure-as-Code
    • Container platforms (Kubernetes, OpenShift)
    • Observability and monitoring platforms
  • Experience with:
    • Incident Management
    • Problem Management
    • Change Management
    • Disaster Recovery
    • Business Continuity
    • Service Reliability Programs
  • Experience leading operational transformations and continuous-improvement initiatives

Preferred Qualifications:
  • Experience in banking, financial services, healthcare, or other highly regulated industries
  • Experience implementing enterprise observability solutions such as Datadog, Splunk, Grafana, Prometheus, or OpenTelemetry
  • Experience with cloud-native architectures and platform engineering practices
  • Experience deploying AIOps, ChatOps, or intelligent automation solutions
  • Familiarity with:
    • Agentic AI workflows
    • LLM-powered operational tooling
    • Knowledge management platforms
    • AI-enabled incident management solutions
  • Experience establishing SLO frameworks and reliability governance programs
  • Leadership Competencies
  • Strategic thinker with solid operational and technology acumen
  • Proven ability to build, lead, and inspire high-performing engineering and operational teams
  • Solid understanding of service reliability, operational risk, resiliency, and governance
  • Exceptional stakeholder management and executive communication skills
  • Data-driven decision maker with a focus on measurable outcomes
  • Solid execution and delivery leadership in complex enterprise environments
  • Collaborative leader focused on continuous improvement, operational excellence, and customer outcomes

*All employees working remotely will be required to adhere to UnitedHealth Group's Telecommuter Policy.
Pay is based on several factors including but not limited to local labor markets, education, work experience, certifications, etc. In addition to your salary, we offer benefits such as, a comprehensive benefits package, incentive and recognition programs, equity stock purchase and 401k contribution (all benefits are subject to eligibility requirements). No matter where or when you begin a career with us, you'll find a far-reaching choice of benefits and incentives. The salary for this role will range from $112,700 to $193,200 annually based on full-time employment. We comply with all minimum wage laws as applicable.
Application Deadline: This will be posted for a minimum of 2 business days or until a sufficient candidate pool has been collected. Job posting may come down early due to volume of applicants.
At UnitedHealth Group, our mission is to help people live healthier lives and make the health system work better for everyone. We believe everyone-of every race, gender, sexuality, age, location and income-deserves the opportunity to live their healthiest life. Today, however, there are still far too many barriers to good health which are disproportionately experienced by people of color, historically marginalized groups and those with lower incomes. We are committed to mitigating our impact on the environment and enabling and delivering equitable care that addresses health disparities and improves health outcomes - an enterprise priority reflected in our mission.
UnitedHealth Group is an Equal Employment Opportunity employer under applicable law and qualified applicants will receive consideration for employment without regard to race, national origin, religion, age, color, sex, sexual orientation, gender identity, disability, or protected veteran status, or any other characteristic protected by local, state, or federal laws, rules, or regulations.
UnitedHealth Group is a drug - free workplace. Candidates are required to pass a drug test before beginning employment.


What UnitedHealth Group employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom