2

Remote Observability Jobs in New York (NOW HIRING)

If you're located beyond that distance, the role is fully remote. For location-specific details ... Integrate observability features with popular frontend tools, frameworks, and build systems to ...

If you're located beyond that distance, the role is fully remote. For location-specific details ... Integrate observability features with popular frontend tools, frameworks, and build systems to ...

About the Role We are seeking a Software Engineer, Security Observability to join our Security team ... This role is open to remote employees, or relocation assistance is available to one of our OpenAI ...

Senior AI Engineer (US Remote)

New York, NY ยท Remote

$180K - $210K/yr

AI observability & monitoring * Responsible AI & safety practices Skills & Experience Required ... REMOTE Basic Requirements * 8+ years experience in software development * AND 2+ years with AI ...

Sr. Platform Engineer (Remote - US)

Manhattan, NY ยท On-site +1

$115K - $157K/yr

Establish observability as a default platform capability by ensuring standardized metrics, logs ... Remote work environment California Applicants: CCPA/CPRA Notice Right to Work Poster Notice of E ...

Senior Software Engineer, Release Infra

New York, NY ยท On-site +1

$192K - $240K/yr

... observability, and incident management processes. You will partner closely with product, platform ... As a perk, we also have up to four weeks per year of fully remote work! Responsibilities * Design ...

next page

Showing results 1-20

Remote Observability information

What are some common challenges faced by professionals in a remote observability role, and how can they be addressed?

Professionals in Remote Observability often face challenges such as monitoring complex, distributed systems, ensuring reliable data collection, and quickly identifying the root causes of issues without physical access to infrastructure. To address these challenges, it's essential to implement robust monitoring tools, establish clear alerting thresholds, and maintain strong communication with development and operations teams. Regular knowledge-sharing sessions and continuous learning about new observability platforms can also help remote teams stay effective and proactive.

What is the difference between Remote Observability vs Remote Monitoring?

AspectRemote ObservabilityRemote Monitoring
FocusComprehensive system insights, including logs, metrics, and tracesTracking specific system metrics and alerts
ToolsOpenTelemetry, Grafana, JaegerNagios, Zabbix, Datadog
Work EnvironmentDevOps, SRE teams managing complex distributed systemsIT operations teams overseeing system health
CredentialsKnowledge of cloud platforms, scripting, and monitoring toolsBasic networking, system administration skills

Remote Observability provides a holistic view of system health through logs, metrics, and traces, enabling proactive troubleshooting. Remote Monitoring focuses on tracking specific metrics and alerts to detect issues. While both roles involve system oversight, observability offers deeper insights for complex environments, whereas monitoring emphasizes real-time alerts for system stability.

What are the key skills and qualifications needed to thrive as a remote observability engineer?

To thrive as a Remote Observability Engineer, you need expertise in monitoring, logging, and tracing, typically supported by experience in systems administration or DevOps and a relevant technical degree. Familiarity with observability tools like Prometheus, Grafana, Datadog, ELK Stack, and cloud monitoring platforms, as well as certifications such as AWS Certified Cloud Practitioner or Google Professional Cloud DevOps Engineer, is highly valued. Strong analytical thinking, problem-solving, and effective communication are vital soft skills for diagnosing issues and collaborating with distributed teams. These skills and qualifications ensure reliable system performance, rapid incident response, and seamless user experiences in complex, cloud-based environments.

What is remote observability?

Remote observability refers to the ability to monitor, measure, and understand the state and performance of systems, applications, or infrastructure from a distance, typically using specialized tools and platforms. It is crucial for organizations that operate distributed or cloud-based environments, as it allows teams to detect issues, analyze metrics, and ensure reliability without needing physical access to the hardware. Remote observability often involves collecting logs, metrics, traces, and other telemetry data to provide a comprehensive view of system health and performance.
What are the most commonly searched types of Observability jobs in New York? The most popular types of Observability jobs in New York are:
What job categories do people searching Remote Observability jobs in New York look for? The top searched job categories for Remote Observability jobs in New York are:
What cities in New York are hiring for Remote Observability jobs? Cities in New York with the most Remote Observability job openings:

Observability SME - Remote

NAVA Software Solutions

Jersey City, NJ โ€ข On-site, Remote

Full-time

Re-posted 2 days ago


Job description

NAVA Software is looking for an Observability SME
Details:
Observability SME
Location: 100% Remote
Duration: 6 -12 months
Job Description
Traceable, event-based Observability enables exploratory investigation when issues occur, for causes both known and unknown. It helps teams troubleshoot issues without first having to predict what or how problems may happen, especially with complex, multi-layer distributed applications connected with microservices. Observability also helps teams improve their understanding of how customers use digital products. Product teams use that awareness to influence future development. They contribute to the overall strategic vision of the organization for Observability capabilities, processes, patterns, and tooling. This role will lead efforts working closely with the Product and Infrastructure teams to ensure that all aspects of the telemetry from applications, business events, appliances and infrastructure are accurately received, tagged, and reported. The role involves leading efforts to maintain observability platform, and ensure it is optimized and operating within SLA's and SLO's. Acting as SME in Observability practices for the enterprise and providing services/solutions across the enterprise that enables businesses to achieve and sustain a higher SLA by improving quality of software, reducing problem determination/down time and over all enhancing the end user experience.
The main responsibilities are:
  • Socializing the Observability capabilities, processes, and Technology with the various application groups.
  • Working with various product and business groups to help determine SLIs, SLOs and SLAs for products, applications, and services offered to the customer. Establishing strategies, processes, and tooling to adhere to the SLAs.
  • Lead efforts to provide self-service capabilities to analyze and visualize Observability data providing End to End visibility to Products and application performance (this will include Dashboards, Alerting, automated incident response capabilities etc,.).
  • Providing strategic roadmap for Observability maturity including recommendations on tooling, capabilities to support the ever-growing enterprise needs and new products.
  • Create, support, and sustain methods and procedures to measure outcomes of Observability practices.
  • Provide ability for developers to use tools to identify symptoms and diagnose application issues by providing them requisite access levels and training
  • Develop and document Observability standards, procedures, and best practices for using the tool, provide education in the tools use.
  • Clearly communicate to IT and business stakeholders regarding performance-related recommendations and tradeoffs.
  • Partner with QA team, assisting with creating and refining effective performance test objectives, test plans, and scenarios that help the organization achieve quality requirements for applications.
  • Work with business to provide guidance for developing KPI's in support of strategize business initiatives.
  • Establish measurements for KPI's and related business transactions of interest and develop executive dashboards required to observe application, user behavior, and user-interaction for business-critical functions.
  • Work with development and architecture teams to manage Observability data collection, analysis, and visualization for critical applications through the lifecycle of the application.
  • Working on continuous improvements of Observability capabilities, providing technical guidance to development teams and aid in triaging production problems
  • Independently utilizes Observability tools to detect, isolate, and resolve issues effecting positive user experience and user interaction with the applications.
  • Assist in major application and/or security incident troubleshooting.
  • Contribute to aspects of the solution delivery lifecycle in prototyping, capacity modeling, performance driven design, profiling, performance testing, availability management, and troubleshooting.
  • Guide operations and support team on building and refining application behavior data capture and reporting for Production systems, and corresponding processes.
  • Provide and design cross-team training opportunities.
  • Improve knowledge and skills in Enterprise Devops team to become more competent and able to accept greater responsibilities.
  • Install and configure software products. Ensure compatibility between target product, operating system, and other resident software. Apply maintenance according to best practices.
  • Lead in capacity planning and performance management activities.
  • Contribute to the development of service level goals and objectives.
  • Develop and prepare metrics that measure services rendered.
  • Identify opportunities to improve service levels and/or minimize support efforts.
  • Perform standard configuration, management, and maintenance tasks in support of web resources.
  • Mentor and/or provide guidance to all members of the team.
  • Participate in disaster planning/mitigation/recovery.
  • Conduct Product Proof-of- Concepts.
  • Assist with other projects as may be required to contribute to the efficiency and effectiveness of the group and other business/technical entities.
  • Assist and participate with Change Management preparations and implementations, providing technical subject matter expertise.
  • Attend, and periodically lead meetings in participation with the team.
  • Participate in hiring activities and fulfilling affirmative action obligations and ensuring compliance with the equal employment opportunity policy.
  • Provide periodic 24/7 on-call support of specific functions.

Required Skills
Specific Skill Set:
  • Software engineering and monitoring tolls like Dynatrace and/or Open Telemetry or any other Opensource tools
  • 10+ years of experienceโ€ข 6-8 years experience Automation experience including CI/CD

Desired Skills:
  • 5+ years of Software Engineering experience
  • 5+ years of experience in Observabilityโ€ข Monitoring software exposure to SRE

NAVA Software Solutions logo

About NAVA Software Solutions

Sourced by ZipRecruiter

NAVA is a strategic partner for companies seeking to develop or customize software and products. Our team of experts leverages cutting-edge technology and deep industry knowledge to provide customized solutions that drive business success. Whether you're looking to improve your operations, increase efficiency, or bring a new product to market, NAVA has the expertise and resources to help you achieve your goals. Trust us to be your partner in software and product development.

Industry

It services

Company size

51 - 200 Employees

Headquarters location

Rocky Hill, CT, US

Social media