1

Network Observability Jobs in New York (NOW HIRING)

Vercel users rely on Observability to monitor and understand their applications' health and ... Learn and Grow - we provide mentorship and send you to events that help you build your network and ...

Vercel users rely on Observability to monitor and understand their applications' health and ... Learn and Grow - we provide mentorship and send you to events that help you build your network and ...

Showing results 21-40

Network Observability information

What is network observability?

Network observability refers to the ability to gain deep visibility into all aspects of a computer network’s operations, performance, and health. It involves collecting and analyzing data from various sources such as logs, metrics, and traces to detect issues, optimize performance, and ensure security. Unlike traditional monitoring, observability provides insights into the underlying causes of network problems, enabling faster troubleshooting and proactive management. Network observability tools help organizations maintain reliable, secure, and efficient network infrastructure, which is especially important in complex or large-scale environments.

What is the difference between Network Observability vs Network Monitoring?

AspectNetwork ObservabilityNetwork Monitoring
FocusProvides comprehensive insights into network performance, health, and security through data collection, analysis, and visualization.Tracks network uptime, availability, and basic performance metrics to detect outages or issues.
Tools & SkillsUses advanced analytics, telemetry, and machine learning; requires knowledge of data analysis and network architecture.Utilizes monitoring tools like SNMP, ping, and simple dashboards; requires basic network troubleshooting skills.
PurposeEnables proactive troubleshooting, capacity planning, and security analysis by understanding complex network behaviors.Provides real-time alerts and status updates to quickly identify and resolve network issues.

While both roles focus on network health, Network Observability offers a deeper, data-driven understanding of network behavior, supporting proactive management. Network Monitoring is more about real-time detection of outages and basic performance tracking. Organizations often use both to ensure robust network performance and security.

What are some common challenges faced by professionals in network observability roles, and how can they be addressed?

Professionals in Network Observability often face challenges like managing large volumes of network data, integrating diverse monitoring tools, and quickly identifying the root cause of network issues. Addressing these challenges typically involves automating data collection, leveraging advanced analytics, and fostering close collaboration with network engineers and security teams. Staying updated with the latest observability platforms and best practices can also help streamline workflows and improve network performance monitoring.

What are the key skills and qualifications needed to thrive in network observability, and why are they important?

To excel in Network Observability, you need a strong understanding of networking fundamentals, troubleshooting, and data analysis, often supported by a degree in computer science or a related field. Familiarity with monitoring tools like Wireshark, Datadog, Grafana, and knowledge of protocols such as SNMP and NetFlow, as well as relevant certifications (e.g., Cisco CCNA/CCNP), are typically required. Strong problem-solving, attention to detail, and communication skills help professionals quickly identify issues and work with cross-functional teams. These competencies are essential for ensuring network reliability, performance, and rapid incident response in complex IT environments.
What are popular job titles related to Network Observability jobs in New York? For Network Observability jobs in New York, the most frequently searched job titles are:
What cities in New York are hiring for Network Observability jobs? Cities in New York with the most Network Observability job openings:
Infographic showing various Network Observability job openings in New York as of July 2026, with employment types broken down into 1% As Needed, 81% Full Time, 13% Part Time, and 5% Contract. Highlights an 93% Physical, 2% Hybrid, and 5% Remote job distribution.

Staff / Principal DevOps Engineer

PVH (Tommy Hilfiger/Calvin Klein)

Manhattan, NY • On-site

$180 - $260/hr

Other

Posted 20 days ago


PVH Corp. rating

6.3

Company rating: 6.3 out of 10

Based on 7 frontline employees who took The Breakroom Quiz


Job description

Staff / Principal DevOps EngineerLocation: New York City | HybridDepartment: AI Platform & Infrastructure TeamReports to: Vangie Shue – Principal Engineering Manager

About AppGate

AppGate secures and protects an organization’s most valuable assets with its high performance Zero Trust Network Access (ZTNA) solution and Cyber Advisory Services. AppGate ZTNA is the only direct‑routed Zero Trust solution built for peak performance, superior protection and seamless interoperability. AppGate Cyber Advisory Services harden your security posture and ensure business continuity. AppGate safeguards Fortune 500 enterprises and government agencies worldwide.

About the Role

As we expand our platform, we are standing up a new AI Platform & Infrastructure team: the engine room of AppGate’s AI strategy. This team owns the infrastructure layer that every next‑generation security capability is built on, from network observability to AI‑driven threat detection and the secure operation of emerging Agentic AI systems. We need a Staff or Principal DevOps Engineer to design and build the deployment vehicle of our cloud and AI infrastructure, set the standards the rest of the team will build on, and grow the surrounding engineers into a capable platform team. You will make foundational decisions around cloud architecture, IaC conventions, CI/CD, observability, and on‑call, codify them so they scale, and mentor a team that is still developing deep DevOps expertise.

Key Responsibilities
  • Build the Platform: design, implement, and maintain CI/CD pipelines and self‑service tooling from a greenfield starting point, enabling product teams to build, test, and deploy safely and quickly with automated guardrails, rollbacks, and paved paths.
  • Stand up and operate container orchestration with Docker and Kubernetes, including self‑managed, stateful technologies such as Kafka and Elasticsearch, running reliably in production.
  • Infrastructure as Code: automate infrastructure provisioning and configuration management using Terraform and Helm, with reusable modules and patterns others can safely build on.
  • Observability from day one: instrument monitoring, logging, and alerting across services and infrastructure using Prometheus, Grafana, and the ELK stack; define SLOs, error budgets, and operational health metrics the team will trust.
  • Security Best Practices: implement security best practices across infrastructure and deployment workflows—including secrets management, least‑privilege access, certificate and key management, vulnerability remediation, and secure‑by‑default configuration.
  • Mentor and level up the team: coach engineers who are strong but new to DevOps/infra, run design reviews, pair on hard problems, write the docs and playbooks, and deliberately transfer expertise so the team can own and evolve the platform without depending on any one person.
  • Operationalize AI/ML: build the MLOps foundation for AppGate’s AI products—model serving and inference pipelines, model deployment, versioning, and drift monitoring, and the experiment‑tracking and lifecycle tooling (MLflow, Kubeflow, SageMaker) that lets data scientists ship models safely and repeatably.
  • Collaborate cross‑functionally: partner with data scientists, product teams, and leadership to align platform investment with AppGate’s strategic vision.
Required Qualifications
  • Extensive DevOps, platform, infrastructure, or SRE engineering experience with a track record of operating production systems at scale (8+ years for Staff, 12+ years for Principal).
  • Hands‑on experience standing up infrastructure and CI/CD pipelines in a greenfield or early‑stage environment and making foundational decisions.
  • Observability expertise implementing Prometheus, Grafana, OpenTelemetry, or the ELK stack, along with SLO and error‑budget practice.
  • Security mindset: familiarity with implementing secrets management, vulnerability remediation, and secure‑by‑default design in infrastructure and deployment workflows.
  • Engineering craft: fluency in a primary backend language (Python, Go, or similar) and a strong bias toward automation, testing, and reliable, maintainable systems.
  • Mentorship & leadership experience: a proven ability to coach and elevate the skills of engineering peers.
Benefits

AppGate offers a comprehensive benefits package, competitive compensation, and opportunities to work on cutting‑edge AI and security solutions in a hybrid, New York City environment.

#J-18808-Ljbffr

What PVH Corp. employees say

Pay

Hours and flexibility

Workplace

Get the full story on Breakroom