1

Sre Jobs in Edison, NJ (NOW HIRING)

Site Reliability Engineer (SRE)

Manhattan, NY ยท On-site

$120K - $180K/yr

You will be the first dedicated Site Reliability Engineer and own critical infrastructure end to end. This is a greenfield opportunity to architect the path from AWS to on-premises and air-gapped ...

next page

Showing results 1-20

Sre information

See Edison, NJ salary details

$11

$65

$95

How much do sre jobs pay per hour?

As of Sep 8, 2026, the average hourly pay for sre in Edison, NJ is $65.99, according to ZipRecruiter salary data. Most workers in this role earn between $56.73 and $75.38 per hour, depending on experience, location, and employer.

What is an SRE?

Site Reliability Engineers (SREs) are IT professionals who use a combination of software engineering and systems administration skills to build and maintain reliable, scalable, and efficient IT infrastructure. SREs are responsible for ensuring that services are available, performant, and resilient by automating operations, monitoring systems, and responding to incidents. Their work focuses on reducing manual processes, improving service reliability, and enabling faster development cycles. SREs often collaborate with development and operations teams to implement best practices and ensure system health.

What skills and qualifications are needed to be an SRE?

To thrive as an SRE, you need a solid background in computer science or engineering, strong programming/scripting abilities, and experience with systems administration and cloud platforms. Familiarity with tools such as Kubernetes, Docker, Prometheus, Terraform, and CI/CD pipelines, along with relevant certifications like AWS Certified Solutions Architect or Google Professional Cloud DevOps Engineer, is highly beneficial. Excellent problem-solving skills, effective communication, and a proactive approach to incident management and collaboration are vital soft skills. These competencies are crucial for maintaining reliable, scalable systems and ensuring rapid incident response in dynamic production environments.

How does an SRE collaborate with development and operations teams to maintain system reliability?

Site Reliability Engineers (SREs) work closely with both development and operations teams to ensure systems are reliable, scalable, and efficient. They often participate in code reviews, incident response, and post-mortem analyses, bridging gaps between software development and IT operations. SREs also help define service-level objectives (SLOs) and implement automation to reduce manual work, fostering a culture of shared responsibility for uptime and performance. Effective communication and cross-team collaboration are central to success in this role.

What is the difference between Sre vs DevOps Engineer?

AspectSreDevOps Engineer
CertificationsOften includes cloud certifications (AWS, GCP), Linux, and scripting skillsSimilar certifications, with focus on automation and cloud platforms
Work EnvironmentFocuses on reliability, monitoring, and incident response in production systemsEmphasizes automation, CI/CD pipelines, and infrastructure management
Industry UsagePrimarily in tech companies with large-scale systems, especially cloud-basedWidely used across tech, startups, and enterprises implementing DevOps practices

Both Sre and DevOps Engineer roles require overlapping skills in automation, cloud platforms, and scripting. Sre emphasizes system reliability and incident management, while DevOps focuses on continuous integration, deployment, and infrastructure automation. The roles often collaborate but have distinct primary focuses within the software development lifecycle.

Is SRE a good career path?

Site Reliability Engineering (SRE) is a growing field that combines software engineering and systems administration to ensure the reliability and availability of large-scale systems. It typically requires skills in programming, automation, and monitoring tools, and offers opportunities for career advancement in tech companies. Many professionals find it a rewarding career due to its focus on system stability and continuous improvement.

What is the average salary of an SRE?

The average salary of a Site Reliability Engineer (SRE) typically ranges from $90,000 to $150,000 annually, depending on experience, location, and company size. SREs with advanced skills in cloud platforms, automation, and monitoring tools tend to earn higher salaries.

What are popular job titles related to Sre jobs in Edison, NJ?

For Sre jobs in Edison, NJ, the most frequently searched job titles are:

What job categories do people searching Sre jobs in Edison, NJ look for?

The top searched job categories for Sre jobs in Edison, NJ are:

What cities near Edison, NJ are hiring for Sre jobs?

Cities near Edison, NJ with the most Sre job openings:

Infographic showing various Sre job openings in Edison, NJ as of August 2026, with employment types broken down into 92% Full Time, 4% Part Time, and 4% Contract. Highlights an 75% Physical, 9% Hybrid, and 16% Remote job distribution, with an average salary of $137,257 per year, or $66 per hour.

Site Reliability Engineer (SRE)

Long Finch Technologies

Kearny, NJ โ€ข On-site

$59.50 - $79.25/hr

Full-time

Posted 27 days ago


Job description

Overview

We are seeking an experienced Site Reliability Engineer (SRE) โ€“ Microsoft Hyper-V & Private Cloud to operate highly available private cloud and Virtual Desktop Infrastructure (VDI) platforms based on Microsoft Hyper-V. This role combines deep Hyper-V expertise with modern Site Reliability Engineering practices to improve platform reliability, scalability, performance, automation, and operational excellence.

The ideal candidate will be responsible for ensuring platform reliability through infrastructure automation, OS upgrades, proactive monitoring, incident response, capacity planning, and continuous service improvement while collaborating with infrastructure, security, and application teams.

Key Responsibilities

  • Operate enterprise-scale private cloud infrastructure built on Microsoft Hyper-V.
  • Optimize, and support highly available VDI environments on Hyper-V.
  • Improve platform reliability, availability, scalability, and resiliency by applying SRE principles and engineering best practices.
  • Disaster recovery, backup, patch management, and business continuity strategies.
  • Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and operational metrics for critical infrastructure services.
  • Automate infrastructure provisioning, configuration management, and operational workflows using PowerShell and Infrastructure as Code (IaC) principles wherever applicable.
  • Manage Hyper-V Failover Clusters, host lifecycle, storage, networking, and capacity to ensure high availability and business continuity.
  • Develop proactive monitoring, alerting, logging, and observability capabilities to detect and prevent service degradation.
  • Lead incident response for infrastructure-related outages, perform root cause analysis (RCA), and implement preventive actions through post-incident reviews.
  • Perform capacity planning, performance tuning, and resource optimization across Hyper-V clusters and VDI platforms.
  • Support infrastructure migration initiatives including P2V, V2V, workload modernization, and private cloud transformations.
  • Collaborate closely with Security, Networking, Platform Engineering, and Application teams to improve platform reliability and operational efficiency.
  • Develop and maintain technical documentation, architecture diagrams, operational runbooks, automation scripts, and standard operating procedures.
  • Mentor junior engineers and promote SRE culture, automation, and operational best practices across the team.


Experience & Qualifications

  • 6+ years of infrastructure engineering experience with at least 4+ years of hands-on Microsoft Hyper-V administration.
  • Demonstrated experience operating mission-critical enterprise infrastructure with high availability and reliability requirements.
  • Proven experience implementing automation to reduce operational overhead and improve service reliability.
  • Experience supporting enterprise private cloud and VDI environments.
  • Experience participating in incident response, root cause analysis, and continuous service improvement initiatives.
  • Microsoft certifications such as Microsoft Certified: Windows Server Hybrid Administrator Associate or equivalent are desirable.
  • Experience in Banking or Financial Services environments is advantageous.

    Preferred Skills

    • Windows Server 2016/2019/2022 administration.
    • Experience with backup and disaster recovery solutions such as Veeam, Altaro, or native Hyper-V Replica.
    • Exposure to hybrid cloud and private cloud platforms.
    • Familiarity with monitoring and observability platforms such as SCOM, Azure Monitor, Prometheus, Grafana, Splunk, or similar tools.
    • Experience supporting enterprise VDI environments.
    • Understanding of ITIL Incident, Problem, Change, and Release Management.
    • Experience working in regulated industries such as Banking or Financial Services.