1

Sre Engineer Devops Engineer Jobs in New York (NOW HIRING)

... a Devops role, strictly need an SRE Engineer, who has great analytical skills and is a good incident manager as well) This is an SRE role supporting the B2B and Core Services. Skills: * SRE. * REST ...

DevOps Engineer

New York, NY · On-site

$57.75 - $79/hr

I am reaching out to share an excellent career opportunity for the role of " Senior Application Cloud SRE/DevOps Engineer" with our esteemed client. If you are interested then please share your ...

SRE/DevOps Engineer

New York, NY · On-site

$62.25 - $82.75/hr

Versana is seeking a motivated SRE/DevOps Engineer with strong observability experience to join our growing Platform Engineering squad. The squad's goal is to manage public cloud, improve DevOps ...

SRE/DevOps Engineer

New York, NY · On-site

$62.25 - $82.75/hr

Versana is seeking a motivated SRE/DevOps Engineer with strong observability experience to join our growing Platform Engineering squad. The squad's goal is to manage public cloud, improve DevOps ...

SRE/DevOps Engineer

New York, NY · On-site

$62.25 - $82.75/hr

Versana is seeking a motivated SRE/DevOps Engineer with strong observability experience to join our growing Platform Engineering squad. The squad's goal is to manage public cloud, improve DevOps ...

Application SRE (DevOps)

Elmwood Park, NJ · On-site

$100K - $120K/yr

We are looking for an Application Site Reliability Engineer (SRE) with strong DevOps experience to improve the reliability, scalability, and performance of our applications. The Application Site ...

AWS SRE

Newark, NJ · On-site

$59.50 - $79/hr

Job Title: Senior AWS Site Reliability Engineer (SRE) Location: New Jersey / Columbus, OH ... AWS Certified DevOps Engineer or AWS Certified Solutions Architect. * Experience with Helm, ArgoCD ...

Senior SRE Engineer

New York, NY · On-site

$62.25 - $82.75/hr

Collaborate closely with development, DevOps, and security teams to design resilient, secure, and ... Preferred, but not required: * SRE Foundation. * AWS Certified Solutions Architect / Azure ...

AI Platform / SRE Lead

OR · Remote

$53.50 - $71/hr

Drive adoption of AI-enabled DevOps and SRE practices * Improve operational efficiency and reliability through automation and AI * Serve as a technical advisor and escalation point across multiple ...

New

AWS DevOps Engineer

New York, NY · On-site

$57.75 - $79/hr

Application Cloud SRE/DevOps Engineer Senior Location - New York, NY (Must be local to the NYC area to be considered) Duration - 12+ Months Bachelor from US required Hybrid work schedule - 3 days ...

next page

Showing results 1-20

Sre Engineer Devops Engineer information

What is an SRE engineer or DevOps engineer?

SRE (Site Reliability Engineer) and DevOps Engineer are roles focused on ensuring the reliability, scalability, and efficiency of software systems. SREs typically use software engineering approaches to automate IT operations tasks, improve system reliability, and manage incidents. DevOps Engineers focus on bridging the gap between development and operations by automating deployments, managing CI/CD pipelines, and fostering a collaborative culture. Both roles share similar goals but differ in focus: SRE emphasizes reliability and incident management, while DevOps centers around streamlining development and operational processes.

What skills and qualifications are needed to thrive as an SRE engineer or DevOps engineer?

To thrive as an SRE Engineer or DevOps Engineer, you need strong knowledge of system administration, cloud infrastructure, automation, and scripting languages, typically backed by a degree in computer science or related experience. Familiarity with tools such as Kubernetes, Docker, Jenkins, Terraform, and monitoring systems, as well as certifications like AWS Certified DevOps Engineer or Google Professional DevOps Engineer, are highly valuable. Excellent problem-solving, collaboration, and communication skills help you work effectively across teams and respond to incidents rapidly. These combined skills and qualities ensure the reliability, scalability, and efficiency of critical production systems in dynamic technology environments.

What are common challenges SRE engineers or DevOps engineers face when balancing reliability with rapid deployment?

SRE/DevOps Engineers often encounter the challenge of maintaining system reliability and uptime while supporting rapid development and frequent deployments. Balancing these priorities requires implementing automated testing, robust monitoring, and efficient rollback strategies to minimize risk. Effective collaboration with development and operations teams is essential to ensure that new features are released safely without compromising system stability. Continuous communication and clear incident response processes also help address issues quickly and maintain service quality.

What is the difference between Sre Engineer Devops Engineer vs Cloud Engineer?

AspectSre Engineer Devops EngineerCloud Engineer
Primary FocusEnsuring system reliability, automation, and performance of servicesDesigning, deploying, and managing cloud infrastructure and services
Skills & CertificationsLinux, scripting, CI/CD, monitoring tools, often certifications like AWS, Azure, GCPCloud platforms (AWS, Azure, GCP), cloud architecture, certifications like AWS Certified Solutions Architect
Work EnvironmentOperations, development, and infrastructure teams in tech companiesCloud service providers, cloud-focused teams, DevOps teams

While Sre Engineers Devops Engineers focus on system reliability, automation, and performance, Cloud Engineers specialize in designing and managing cloud infrastructure. Both roles often collaborate but have distinct core responsibilities within tech organizations.

Are SRE and DevOps engineers the same?

SRE (Site Reliability Engineer) and DevOps engineers both focus on improving software deployment and system reliability, but their roles differ; SREs typically emphasize reliability, monitoring, and automation using tools like Prometheus and Kubernetes, while DevOps engineers focus on integrating development and operations processes to enable continuous delivery. Both roles require strong scripting skills and collaboration across teams, but SREs often have a more specialized focus on system stability and incident response.

Site Reliability Engineer

Wood Ridge, NJ • On-site

PROLIM Global Corporation
IT Services • 201 - 500 employees

$58 - $77.25/hr

Contractor

Re-posted 5 days ago


Job description

Job Title: Site Reliability Engineer (SRE)
Location: Wood Ridge, NJ
Duration: 6 Months
Position Overview
We are seeking a highly skilled Site Reliability Engineer (SRE) to join our digital engineering and operations team. The ideal candidate will bring a blend of SRE best practices, DevOps mindset, and IBM WebSphere Commerce Suite (WCS) expertise to ensure the reliability, scalability, and performance of our digital platforms.
This role is responsible for building and maintaining automated systems that monitor, measure, and enhance service uptime and performance. The engineer will collaborate closely with development, infrastructure, and operations teams to drive continuous improvement across the environment.
Key Responsibilities
  • Design, implement, and maintain scalable, reliable, and highly available systems to support critical business applications.
  • Manage and optimize IBM WebSphere Commerce Suite (WCS) environments for high performance and fault tolerance.
  • Develop and implement automation tools and scripts for deployment, monitoring, and maintenance tasks.
  • Proactively monitor system performance, identify bottlenecks, and recommend improvements to increase efficiency and resiliency.
  • Collaborate with software development teams to define service-level objectives (SLOs) and ensure service-level indicators (SLIs) meet operational goals.
  • Perform root cause analysis of production issues and drive permanent corrective actions.
  • Implement and manage DevOps pipelines for continuous integration and deployment (CI/CD).
  • Use monitoring tools (e.g., Prometheus, Grafana, ELK, Splunk, AppDynamics, New Relic) to maintain service health visibility.
  • Apply chaos engineering principles to test system resilience and recovery.
  • Participate in on-call rotations and incident response, ensuring quick restoration of service.
Required Skills & Qualifications
  • 5+ years of hands-on experience as a Site Reliability Engineer, DevOps Engineer, or Systems Engineer.
  • Proven experience with IBM WebSphere Commerce Suite (WCS) deployment, configuration, and optimization.
  • Strong knowledge of SRE principles - monitoring, automation, incident management, reliability metrics (SLO/SLI).
  • Proficiency in scripting languages such as Python, Shell, or Groovy.
  • Experience with CI/CD tools (Jenkins, GitLab CI, or similar).
  • Familiarity with containerization technologies (Docker, Kubernetes).
  • Hands-on experience with cloud platforms (AWS, Azure, or GCP).
  • Excellent troubleshooting and problem-solving skills.
  • Strong collaboration and communication abilities with cross-functional teams.
Preferred Qualifications
  • Experience implementing chaos engineering and resiliency testing frameworks.
  • Background in application performance monitoring (APM) and tuning.
  • Exposure to infrastructure-as-code tools (Terraform, Ansible, or CloudFormation).
  • Certification in AWS/Azure/GCP or related SRE/DevOps technologies.