1

Home Based Observability Engineer Jobs in California

The Observability Engineering organization at CoreWeave is responsible for the platforms and ... The starting salary will be determined based on job-related knowledge, skills, experience, and ...

$130K - $171K/yr

Familiarity with OpenTelemetry, distributed systems, eventing, or cloud-based observability ... Fully remote, work from home environment * Employee Share Option Plan * Flexible working hours

The Observability Engineering organization at CoreWeave is responsible for the platforms and ... The starting salary will be determined based on job-related knowledge, skills, experience, and ...

The Observability Engineering organization at CoreWeave is responsible for the platforms and ... The starting salary will be determined based on job-related knowledge, skills, experience, and ...

next page

Showing results 1-20

Home Based Observability Engineer information

What is the difference between Home Based Observability Engineer vs Network Operations Center (NOC) Technician?

AspectHome Based Observability EngineerNetwork Operations Center (NOC) Technician
CredentialsRelevant certifications like Cisco CCNA, CompTIA Network+Similar certifications often required, such as Cisco CCNA
Work EnvironmentRemote, home-based setup with monitoring toolsOn-site or remote, monitoring network systems in a control room
Industry UsageIT, cloud services, software companiesTelecommunications, internet service providers
Job FocusMonitoring, troubleshooting, and optimizing observability toolsNetwork monitoring, incident response, and system maintenance

While both roles involve monitoring network and system health, the Home Based Observability Engineer focuses on software observability tools and cloud environments remotely, whereas the NOC Technician primarily manages network infrastructure on-site or remotely. Both require similar certifications and are vital in maintaining system uptime, but their daily tasks and work settings differ.

What are the most commonly searched types of Observability Engineer jobs in California?

The most popular types of Observability Engineer jobs in California are:

What are popular job titles related to Home Based Observability Engineer jobs in California?

For Home Based Observability Engineer jobs in California, the most frequently searched job titles are:

What job categories do people searching Home Based Observability Engineer jobs in California look for?

The top searched job categories for Home Based Observability Engineer jobs in California are:

What cities in California are hiring for Home Based Observability Engineer jobs?

Cities in California with the most Home Based Observability Engineer job openings:

Senior AI and HPC Observability Engineer

Nvidia Corporation

Santa Clara, CA • On-site

$143K - $189K/yr

Full-time

Re-posted 15 days ago


Nvidia rating

9.6

Company rating: 9.6 out of 10

Based on 17 frontline employees who took The Breakroom Quiz

7th of 244 rated software companies


Job description

NVIDIA is a pioneer in accelerated computing, known for inventing the GPU and driving breakthroughs in gaming, computer graphics, high-performance computing, and artificial intelligence. Our technology powers everything from generative AI to autonomous systems, and we continue to shape the future of computing through innovation and collaboration. Within this mission, our team, Managed AI Superclusters (MARS) builds and scales the infrastructure, platforms, and tools that enable researchers and engineers to develop the next generation of AI/ML systems. By joining us, you'll help design solutions that power some of the world's most advanced computing workloads.
Observability is at the heart of this transformation. We are looking for a strong AI & HPC Observability Engineer to build and scale next-generation Observability and Telemetry platforms. You will design and develop high-throughput, reliable telemetry pipelines and modern data infrastructure. This role requires solid distributed systems fundamentals, production-grade coding, and a passion for operational excellence.
What You Will Be Doing:
  • Design and scale observability platforms handling high-volume metrics, logs, and traces across distributed environments
  • Build high-performance backend services for telemetry ingestion, processing, and routing
  • Develop and extend OpenTelemetry collectors, processors, exporters, and instrumentation libraries
  • Build and optimize metrics pipelines using large-scale time-series storage systems
  • Design and operate real-time and batch telemetry pipelines using streaming and distributed data technologies
  • Improve platform reliability, performance, and cost efficiency through tuning, capacity planning, and system optimization
  • Develop monitoring, alerting, and service reliability frameworks to ensure platform health and performance
  • Collaborate with platform engineering, infrastructure, and site reliability teams to deliver production-grade observability solutions

What We Need to see:
  • Bachelor's degree in Computer Science, Computer Engineering, or related field or equivalent experience
  • 5+ years of experience building backend or distributed systems in production environments
  • Strong programming skills in Python, Go, or Java, with experience developing production-quality software
  • Hands-on experience with modern observability architectures, including metrics, logs, and traces
  • Solid experience with PromQL and time-series data systems
  • Experience building or operating distributed data pipelines using technologies such as Kafka, Spark, or Flink
  • Experience working with Kubernetes and cloud-native infrastructure
  • Strong understanding of distributed systems, concurrency, and fault-tolerant system design. Strong debugging, performance tuning, and production operations skills

Ways To Stand Out from The Crowd:
  • Proven experience designing and scaling observability platforms for AI, GPU, or HPC environments
  • Hands-on expertise with OpenTelemetry, Prometheus, Kafka, and high-volume distributed telemetry pipelines
  • Strong background in data engineering, time-series data modeling, and real-time performance tuning
  • Experience integrating observability with AI/ML pipelines, GPU workload monitoring, or intelligent alerting
  • Demonstrated use of statistical or machine learning techniques for anomaly detection, correlation, or predictive insights

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 152,000 USD - 241,500 USD for Level 3, and 184,000 USD - 287,500 USD for Level 4.
You will also be eligible for equity and benefits.
Applications for this job will be accepted at least until March 6, 2026.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

What Nvidia employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Nvidia logo

About Nvidia

Sourced by ZipRecruiter

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.

Industry

Computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

Santa Clara, CA, US

Year founded

1993