ABOUT THE ROLE We''re hiring a Principal/Staff Observability Platform Engineer to own the technical direction of our observability platform -- the systems that give us deep visibility into GPU ...
ABOUT THE ROLE We''re hiring a Principal/Staff Observability Platform Engineer to own the technical direction of our observability platform -- the systems that give us deep visibility into GPU ...
Staff Software Engineer, Observability
Menlo Park, CA · On-site
$210 - $290/hr
The Observability team's mission is to build and own Robinhood's full-stack observability platform -- the foundation that keeps every product, service, and customer experience running reliably at ...
Staff Software Engineer, Observability
Menlo Park, CA · On-site
$210 - $290/hr
The Observability team's mission is to build and own Robinhood's full-stack observability platform -- the foundation that keeps every product, service, and customer experience running reliably at ...
Staff Software Engineer, Observability
Menlo Park, CA · On-site
$190 - $280/hr
The Observability team's mission is to build and own Robinhood's full-stack observability platform -- the foundation that keeps every product, service, and customer experience running reliably at ...
Staff Software Engineer, Observability
Menlo Park, CA · On-site
$190 - $280/hr
The Observability team's mission is to build and own Robinhood's full-stack observability platform -- the foundation that keeps every product, service, and customer experience running reliably at ...
Engineering Manager Observability
Sunnyvale, CA · On-site
$218K - $335K/yr
This person will start delivering impact through observability frameworks and will evolve depending on business needs but they will be expected to identify high ROI investments with minimal guidance.
Engineering Manager Observability
Sunnyvale, CA · On-site
$218K - $335K/yr
This person will start delivering impact through observability frameworks and will evolve depending on business needs but they will be expected to identify high ROI investments with minimal guidance.
Software Engineer - Observability
San Francisco, CA · On-site
$180 - $240/hr
As an early member of the Observability Team, you will be pivotal in building and shaping the observability experience for our internal and external customers. By joining this team, you'll have a ...
Software Engineer - Observability
San Francisco, CA · On-site
$180 - $240/hr
As an early member of the Observability Team, you will be pivotal in building and shaping the observability experience for our internal and external customers. By joining this team, you'll have a ...
Software Engineer - Observability
San Francisco, CA · On-site
$130 - $180/hr
As an early member of the Observability Team, you will be pivotal in building and shaping the observability experience for our internal and external customers. By joining this team, you'll have a ...
Software Engineer - Observability
San Francisco, CA · On-site
$130 - $180/hr
As an early member of the Observability Team, you will be pivotal in building and shaping the observability experience for our internal and external customers. By joining this team, you'll have a ...
Observability & Monitoring Engineer
Rancho Cucamonga, CA · On-site
$54 - $73.75/hr
Observability & Monitoring Engineer w/d Solarwinds and DynaTrace Exp Location: Rancho Cucamonga, CA 5 Days Onsite Role Duration: Long Term Project (Only Visa independent candidate) Mandatory Skills ...
Quick apply
Observability & Monitoring Engineer
Rancho Cucamonga, CA · On-site
$54 - $73.75/hr
Observability & Monitoring Engineer w/d Solarwinds and DynaTrace Exp Location: Rancho Cucamonga, CA 5 Days Onsite Role Duration: Long Term Project (Only Visa independent candidate) Mandatory Skills ...
As a Staff Engineer on the Observability team, you'll be responsible for designing and building the infrastructure and tools that provide visibility into Pinterest's large-scale distributed systems ...
As a Staff Engineer on the Observability team, you'll be responsible for designing and building the infrastructure and tools that provide visibility into Pinterest's large-scale distributed systems ...
Staff Software Engineer, Observability
Menlo Park, CA · On-site
$180 - $250/hr
The Observability team's mission is to build and own the company's full-stack observability platform - the foundation that keeps every product, service, and customer experience running reliably at ...
Staff Software Engineer, Observability
Menlo Park, CA · On-site
$180 - $250/hr
The Observability team's mission is to build and own the company's full-stack observability platform - the foundation that keeps every product, service, and customer experience running reliably at ...
The Observability team's mission is to build and own Robinhood's full-stack observability platform - the foundation that keeps every product, service, and customer experience running reliably at ...
The Observability team's mission is to build and own Robinhood's full-stack observability platform - the foundation that keeps every product, service, and customer experience running reliably at ...
Senior Manager, Observability
Sunnyvale, CA · On-site
$188K - $275K/yr
The Observability Engineering organization at CoreWeave is responsible for the platforms and practices that help engineers understand, operate, and improve production systems at scale. This team owns ...
Senior Manager, Observability
Sunnyvale, CA · On-site
$188K - $275K/yr
The Observability Engineering organization at CoreWeave is responsible for the platforms and practices that help engineers understand, operate, and improve production systems at scale. This team owns ...
Staff Software Engineer, Observability
San Francisco, CA · On-site
$177.18 - $364.80/hr
As a Staff Engineer on the Observability team, you'll be responsible for designing and building the infrastructure and tools that provide visibility into Pinterest's large‑scale distributed systems ...
Staff Software Engineer, Observability
San Francisco, CA · On-site
$177.18 - $364.80/hr
As a Staff Engineer on the Observability team, you'll be responsible for designing and building the infrastructure and tools that provide visibility into Pinterest's large‑scale distributed systems ...
The role involves building the observability product for OpenAI, focusing on scalable infrastructure and AI-powered tools to enhance system reliability and performance. Responsibilities : • Own ...
The role involves building the observability product for OpenAI, focusing on scalable infrastructure and AI-powered tools to enhance system reliability and performance. Responsibilities : • Own ...
Software Engineer - Observability
San Francisco, CA · On-site
$180 - $260/hr
As an early member of the Observability Team, you will be pivotal in building and shaping the observability experience for our internal and external customers. By joining this team, you'll have a ...
Software Engineer - Observability
San Francisco, CA · On-site
$180 - $260/hr
As an early member of the Observability Team, you will be pivotal in building and shaping the observability experience for our internal and external customers. By joining this team, you'll have a ...
Observability Engineer
Westlake Village, CA · On-site
$61.25 - $84/hr
Job Title: Sr Observability Engineer Location: US-CA-Westlake Village(Onsite)--Local only Long Term Contract Client interview : In-person is mandatory Visa Independent only (USC , GC, GC EAD ...
Quick apply
Observability Engineer
Westlake Village, CA · On-site
$61.25 - $84/hr
Job Title: Sr Observability Engineer Location: US-CA-Westlake Village(Onsite)--Local only Long Term Contract Client interview : In-person is mandatory Visa Independent only (USC , GC, GC EAD ...
Senior Software Engineer, Observability
San Francisco, CA · On-site +1
$200K - $280K/yr
The storage and observability team is crucial for designing, implementing, and maintaining robust distributed storage solutions, ensuring seamless data access and management. They are also ...
Senior Software Engineer, Observability
San Francisco, CA · On-site +1
$200K - $280K/yr
The storage and observability team is crucial for designing, implementing, and maintaining robust distributed storage solutions, ensuring seamless data access and management. They are also ...
Engineering Manager, Observability Infrastructure
San Mateo, CA · On-site
$122K - $160K/yr
The Observability team builds the infrastructure that empowers engineers to understand, operate, and improve the Roblox platform and ecosystem. Our team owns the end-to-end observability stack across ...
Engineering Manager, Observability Infrastructure
San Mateo, CA · On-site
$122K - $160K/yr
The Observability team builds the infrastructure that empowers engineers to understand, operate, and improve the Roblox platform and ecosystem. Our team owns the end-to-end observability stack across ...
As a Staff Engineer on the Observability team, you'll be responsible for designing and building the infrastructure and tools that provide visibility into Pinterest's large-scale distributed systems ...
As a Staff Engineer on the Observability team, you'll be responsible for designing and building the infrastructure and tools that provide visibility into Pinterest's large-scale distributed systems ...
Senior Software Engineer, Observability
San Francisco, CA · On-site
$200K - $280K/yr
The storage and observability team is crucial for designing, implementing, and maintaining robust distributed storage solutions, ensuring seamless data access and management. They are also ...
Senior Software Engineer, Observability
San Francisco, CA · On-site
$200K - $280K/yr
The storage and observability team is crucial for designing, implementing, and maintaining robust distributed storage solutions, ensuring seamless data access and management. They are also ...
Principal Platform Engineer, Observability (CIPE)
Santa Clara, CA · On-site
$147 - $237.50/hr
Responsibilities Observability Architecture * Design and lead the evolution of a modern observability platform using OpenTelemetry, Prometheus, Jaeger, Alertmanager, and related CNCF ecosystem tools.
Principal Platform Engineer, Observability (CIPE)
Santa Clara, CA · On-site
$147 - $237.50/hr
Responsibilities Observability Architecture * Design and lead the evolution of a modern observability platform using OpenTelemetry, Prometheus, Jaeger, Alertmanager, and related CNCF ecosystem tools.
Observability information
See California salary details
$16.13 - $22.41
0% of jobs
$22.41 - $28.68
0% of jobs
$28.68 - $34.96
2% of jobs
$34.96 - $41.24
5% of jobs
$41.24 - $47.51
10% of jobs
$50.45 is the 25th percentile. Wages below this are outliers.
$47.51 - $53.79
17% of jobs
The median wage is $58.74 / hr.
$53.79 - $60.06
20% of jobs
$60.06 - $66.34
18% of jobs
$67.46 is the 75th percentile. Wages above this are outliers.
$66.34 - $72.62
15% of jobs
$72.62 - $78.89
9% of jobs
$78.89 - $85.17
4% of jobs
$16
$59
$85
How much do observability jobs pay per hour?
What is an observability?
An Observability job focuses on ensuring the performance, reliability, and health of software systems by collecting, analyzing, and visualizing telemetry data such as logs, metrics, and traces. Professionals in this field work with monitoring tools, distributed tracing, and alerting systems to detect and troubleshoot issues proactively. They collaborate with engineering and operations teams to improve system visibility, reduce downtime, and enhance overall system performance.
What does an observability do?
In an Observability role, your daily tasks often include designing and maintaining monitoring dashboards, configuring alerts, analyzing system logs, and working closely with development and operations teams to troubleshoot issues. You'll proactively identify areas of improvement to increase system reliability, document monitoring strategies, and support incident response efforts. Collaboration is key, as you may participate in post-incident reviews and help drive architectural improvements based on the data you collect. The role is dynamic and requires a proactive approach to ensure systems stay healthy and downtime is minimized.
What are the key skills and qualifications needed to thrive in an observability role?
To thrive in an Observability role, you need a strong background in monitoring, alerting, logging, and analyzing system performance, often supported by a degree in computer science or related field. Familiarity with tools such as Prometheus, Grafana, Datadog, Splunk, and experience with cloud platforms and scripting languages is crucial. Excellent problem-solving, communication, and collaboration skills help you work effectively with cross-functional engineering and operations teams. These capabilities are essential to ensure system reliability, quickly detect issues, and maintain seamless digital experiences.
Is observability a good career?
What are the most commonly searched types of Observability jobs in California?
The most popular types of Observability jobs in California are:
What are popular job titles related to Observability jobs in California?
For Observability jobs in California, the most frequently searched job titles are:
What job categories do people searching Observability jobs in California look for?
The top searched job categories for Observability jobs in California are:
What cities in California are hiring for Observability jobs?
Cities in California with the most Observability job openings:

Other
Posted 9 days ago
Job description
ABOUT THE ROLE
We''re hiring a Principal/Staff Observability Platform Engineer to own the technical direction of our observability platform — the systems that give us deep visibility into GPU clusters, AI workloads, and the infrastructure running them.
This is a "define, build, and lead" role, not a "maintain and operate" role. You''ll set the architectural roadmap, raise the engineering bar across teams, and make sure the platform scales ahead of the business, not behind it. We have a strong bias toward simplicity — the systems you build should be easy to operate and self-evidently correct when something goes wrong.
RESPONSIBILITIES
- Own the technical strategy and architecture for observability across metrics, logs, traces, and alerting at scale
- Drive platform decisions with multi-year impact: tooling, data models, ingestion patterns, retention, cardinality management
- Identify systemic gaps before they become incidents; design platforms that make failure visible and fast to diagnose
- Partner with SRE, infrastructure, and AI/ML teams to embed observability natively into how we build and operate
- Define standards and patterns that other engineers adopt because they''re clearly better, not by mandate
- Mentor and technically grow the observability team
- Lead incident postmortems and drive durable platform improvements
- Evaluate and introduce tooling that improves signal quality, operational efficiency, or scalability — and retire what doesn''t
REQUIRED SKILLS & EXPERIENCE
- 8+ years in SRE, infrastructure engineering, platform engineering, or observability-focused roles
- Proven experience operating observability infrastructure at serious scale
- Deep hands-on experience with a significant subset of: Prometheus, Thanos, VictoriaMetrics, Grafana, Loki, Tempo, OpenTelemetry, ClickHouse, Elastic
- Strong engineering fundamentals in Python, Go, or similar
- Kubernetes at scale
- Infrastructure-as-Code as default practice (Terraform, Ansible, or equivalent)
- Demonstrated ability to architect systems, write code, review others'' work, and clearly explain tradeoffs
- Track record of influencing engineering direction across teams without formal authority
PREFERRED SKILLS & EXPERIENCE
- Experience with high-volume streaming pipelines for observability data (Kafka, Vector, Fluent Bit, etc.)
- Background in AI/ML infrastructure observability: GPU utilization, training job visibility, inference latency
- Familiarity with GPU infrastructure or HPC environments (Slurm)
- Prior experience defining observability strategy at an organizational level
EQUAL OPPORTUNITY
We strongly encourage applications from people of color, the LGBTQ+ community, people with disabilities, neurodivergent individuals, parents, carers, and people from lower socio-economic backgrounds. If there''s anything we can do to accommodate your specific situation, please let us know.