1

Senior Observability Engineer Jobs in New Mexico

Senior AI OPS Engineer

Albuquerque, NM ยท On-site

$103 - $155/hr

Job Summary: JCS Solutions LLCis seeking a Senior AIOps Engineer to support critical mission ... Missionโ€‘Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified ...

Senior Security Engineer, IAM

Albuquerque, NM ยท On-site

$111K - $152K/yr

Posting Type Remote Job Overview The Senior IAM Engineer is a technically authoritative leader who ... observability. * Working command of industry-standard security benchmarks and frameworks (MITRE ...

Senior AI Systems Engineer

Albuquerque, NM ยท On-site +1

$95K - $130K/yr

... platform engineering teams. * Operationalize machine learning workflows and support AI-enabled ... Maintain observability across AI systems through logging, metrics, performance monitoring, alerting ...

Senior Site Reliability Engineer

Albuquerque, NM ยท On-site

$52 - $69.25/hr

Partner with software developers, platform engineers, and IT staff to improve system design ... Strong experience designing, maintaining, and maturing observability tooling including monitoring ...

$149.58 - $219.38/hr

In this role, you will work directly with engineering, infrastructure, and the CloudOps team to ... Establish standard observability for metrics, logs, traces, and health signals across the platform.

New

Senior Observability Engineer information

What is a senior observability engineer?

A Senior Observability Engineer is a seasoned IT professional responsible for designing, implementing, and maintaining systems that monitor and provide insights into the performance, health, and reliability of software applications and infrastructure. They utilize tools for logging, monitoring, tracing, and alerting to ensure that systems are observable and any issues can be quickly detected and resolved. In addition to technical expertise, they often collaborate with development and operations teams to establish best practices, improve incident response, and optimize system performance. Their work is crucial for maintaining uptime, enhancing customer experiences, and supporting the scalability of technology platforms.

How does a senior observability engineer typically collaborate with development and operations teams?

A Senior Observability Engineer works closely with both development and operations teams to ensure robust monitoring, logging, and tracing solutions are in place across all applications and infrastructure. They often participate in architecture discussions to advise on best practices for instrumenting code and systems for observability. By analyzing metrics and alerting patterns, they help teams proactively resolve issues and optimize system performance. This role also involves mentoring engineers on observability tools and fostering a culture of transparency and accountability in incident response.

What are the key skills and qualifications needed to thrive as a senior observability engineer, and why are they important?

To thrive as a Senior Observability Engineer, you need expertise in monitoring, logging, and tracing systems, with a solid background in computer science or a related field. Familiarity with tools like Prometheus, Grafana, ELK stack, and cloud platforms, as well as certifications such as AWS Certified DevOps Engineer, are typically required. Strong problem-solving, collaboration, and communication skills are critical for effectively diagnosing and resolving complex infrastructure issues. These skills ensure reliable system performance, rapid incident response, and continuous improvement of the technology environment.

What is the difference between Senior Observability Engineer vs Site Reliability Engineer?

AspectSenior Observability EngineerSite Reliability Engineer
CredentialsExperience with monitoring tools, scripting, cloud platformsSame as Senior Observability Engineer, often with SRE certifications
Work EnvironmentFocus on monitoring, logging, and tracing systemsFocus on system reliability, automation, and incident response
Industry UsageUsed in tech companies emphasizing system observabilityCommon in large-scale tech and cloud services
Search/Comparison IntentOften compared for monitoring rolesCompared for reliability and system stability roles

While both roles require expertise in cloud platforms and scripting, the Senior Observability Engineer primarily focuses on designing and maintaining monitoring, logging, and tracing systems to ensure system visibility. In contrast, a Site Reliability Engineer emphasizes system reliability, automation, and incident management to maintain service uptime. Both roles are vital in tech environments but serve different core functions related to system health and stability.

How much do senior observability engineers make?

Senior observability engineers typically earn between $110,000 and $160,000 annually, depending on experience, location, and company size. They often work with tools like Prometheus, Grafana, and cloud platforms, and may require advanced knowledge of monitoring, logging, and alerting systems.

What does a senior observability engineer do?

A senior observability engineer designs, implements, and maintains systems to monitor the performance and health of software applications and infrastructure. They utilize tools like Prometheus, Grafana, and ELK stack to analyze metrics, logs, and traces, ensuring system reliability and performance. This role often requires strong scripting skills and knowledge of cloud environments and distributed systems.

What are the most commonly searched types of Observability Engineer jobs in New Mexico?

The most popular types of Observability Engineer jobs in New Mexico are:

What are popular job titles related to Senior Observability Engineer jobs in New Mexico?

For Senior Observability Engineer jobs in New Mexico, the most frequently searched job titles are:

What job categories do people searching Senior Observability Engineer jobs in New Mexico look for?

The top searched job categories for Senior Observability Engineer jobs in New Mexico are:

What cities in New Mexico are hiring for Senior Observability Engineer jobs?

Cities in New Mexico with the most Senior Observability Engineer job openings:

Senior AI OPS Engineer

JCS Solutions LLC

Kirtland Air Force Base, NM โ€ข On-site

$103K - $155K/yr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

This job post hasย expired today.ย Applications are no longer accepted.


Job description

Grow, innovate, and generate progress: Harness your expertise to solve challenges and celebrate success!
Job Summary:
JCS Solutions LLC is seeking a Senior AIOps Engineer to support critical mission operations within a secure environment and lead the transformation of our IT Service Management (ITSM) capabilities. This role is responsible for the design, deployment, and management of AIOps solutions that enhance the reliability and security of Department of War (DoW) networks and systems.
Acting as the technical lead for this initiative, you will orchestrate integrations across existing Network Engineering, ServiceNow, and SolarWinds teams. You will utilize Splunk and the Machine Learning Toolkit (MLTK) to provide descriptive and predictive analytics and establish closed-loop automated incident response, ensuring the high availability of mission-essential infrastructure.
What's in it for you:
  • Join a premier technology firm specializing in innovative solutions.
  • Be part of a collaborative, inclusive, and innovative work culture.
  • Enjoy tremendous growth potential in a high-performing team environment.
  • A robust benefits package:
    • Health, dental, and vision insurance
    • Life insurance
    • Short-and-long term disability
    • Paid time off (PTO)
    • 401k retirement plan with employer match
    • Annual Professional Development Reimbursement Program
    • And more!
What you will do:
  • Cross-Functional Leadership: Lead the AIOps platform initiative by acting as the primary technical liaison to existing Network Engineering, ServiceNow, and SolarWinds administration teams to establish unified telemetry pipelines.
  • ITSM Orchestration & Automation: Architect closed-loop remediation workflows by deeply integrating Splunk ITSI alerts with ServiceNow Event Management and Incident Management modules.
  • Mission-Critical Observability: Architect and maintain Splunk AIOps solutions across unclassified and classified enclaves to provide real-time situational awareness.
  • Infrastructure Telemetry Integration: Normalize and correlate network performance and fault data from SolarWinds with server and application logs to provide a holistic view of enterprise health.
  • Advanced ML Development: Deploy custom machine learning models via Splunk MLTK to identify anomalous behavior, potential cyber threats, and infrastructure degradations.
  • Secure Data Integration: Engineer secure data ingestion pipelines for telemetry data from cross-domain solutions and tactical edge devices.
  • Incident Reduction: Utilize IT Service Intelligence (ITSI) to correlate multi-source events, reducing noise and prioritizing high-impact mission alerts.
  • Cyber Defense Support: Collaborate with the Cyber Security Service Provider (CSSP) to integrate AIOps insights into defensive cyber operations (DCO).
  • Compliance & Documentation: Ensure all observability tools comply with DoW STIGs and IL5/IL6 protocols; develop and maintain architectural documentation and compliance traceability.
  • Mission Alignment: Stay current on AIOps and related capabilities relevant to DoD, federal, and intelligence mission systems.
What you will bring:
  • Security Clearance: Active Top Secret / Sensitive Compartmented Information (TS/SCI) required at time of hire.
  • Certification: Active IAT Level II certification (e.g., Security+ CE, CySA+, GSEC, or SSCP) required.
  • Citizenship: United States Citizenship is required.
  • Platform Experience: 7+ years of experience with Splunk Enterprise, including architectural design, cluster management, and advanced Search Processing Language (SPL).
  • AIOps & ITSM: 3+ years of experience implementing AIOps workflows, including integration with enterprise ITSM solutions (ServiceNow) for automated root cause analysis and remediation.
  • Machine Learning: Proven track record of building, testing, and tuning supervised and unsupervised models within the Splunk MLTK.
  • Scripting & Automation: Advanced scripting skills for developing custom search commands, API integrations, and automating remediation tasks (e.g., Python).
  • Leadership: Experience leading technical working groups and directing the efforts of adjacent infrastructure and development teams.
  • Operational Experience: Prior experience working within a DoW/DoD Operations Center (NOC/SOC) or supporting mission-critical systems and networks.
  • Communication: Must be able to present designs, plans, and analyses of alternatives to technical leadership boards for approvals.
How you will wow us:
  • Enterprise Aggregation: Experience aggregating and correlating telemetry from diverse tools, specifically SolarWinds, ServiceNow, and VMware vCenter.
  • Expert Certification: Splunk Enterprise Certified Architect or Splunk ITSI Certified Admin.
  • Cloud Observability: Experience with Cloud Native Computing Foundation (CNCF) observability tools in secure hybrid multi-cloud environments (Azure/AWS).
  • RMF/ATO Knowledge: Understanding of the Risk Management Framework (RMF) and the Authorization to Operate (ATO) process for AI/ML workloads.
JCS Solutions (JCS) is a premier technology firm providing innovative solutions and high-quality services in defense, national security, and civilian sectors. JCS offers enterprise-wide solutions including cloud computing, software development, cybersecurity, digital modernization, and management consulting for the federal government. At JCS, we elevate our customers' mission through the application of technology and professional services. Our commitment to investing in our workforce drives innovation and progress for our clients, employees, and communities.
JCS is both a Great Place to Work and a Top Places to Work certified company.
Our employees embody our core values, and we are looking for others who do too!
  • Customer Experience: Strive for excellence and delight our clients
  • Innovation: Embrace creative thinking to enable continual growth and powerful solutions
  • Accountability: Take ownership of and pride in our actions and service delivery
  • Inspire: Be inspired to be your best self and have fun in the process
  • Integrity: Do the right thing, the right way, every time!
  • Stewardship: The careful and responsible management of something entrusted to our care.
At JCS Solutions, compensation is based on a number of factors such as location, qualifications, and applicable contract terms. The general salary range for this position is as follows: $103,000.00 - $155,000.00.
Commitment to non-discrimination: All qualified applicants will receive consideration for employment without regard to any status protected by applicable federal, state, or local laws.
NOTICE: Please be aware that all JCS Solutions communications related to job interviews and offers from our recruiting team willonlycome from @JCSSolutions.com. We want to emphasize thatwe do not conduct any interviews over Discord, Slack, Skype, Zoom, or any chat app.