1

Observability Kubernetes Jobs in California (NOW HIRING)

Create observability and monitoring systems with Prometheus and Grafana to surface cluster health ... Design and implement federated Kubernetes architectures for multi-cluster management. * Build ...

Create observability and monitoring systems with Prometheus and Grafana to surface cluster health ... Design and implement federated Kubernetes architectures for multi-cluster management. * Build ...

Build observability and monitoring systems with Prometheus and Grafana to surface cluster health ... Design and implement federated Kubernetes architectures for multi-cluster management. * Build ...

Build observability and monitoring systems with Prometheus and Grafana to surface cluster health ... Design and implement federated Kubernetes architectures for multi-cluster management. * Build ...

Java Kubernetes Admin

Sunnyvale, CA · On-site

$59.75 - $77.50/hr

You will help validate dashboards and alerts, run and monitor migration scripts, and support the observability team's day-to-day operations. Primary Skill: AWS Services, Kubernetes (Administration ...

Software Engineer - Observability

Palo Alto, CA

$180K - $440K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

Experience operating observability pipelines in Kubernetes or similar orchestration environments. COMPENSATION AND BENEFITS: $180,000 - $440,000 USD Base salary is just one part of our total rewards ...

Engineering Manager Observability

Sunnyvale, CA · On-site

$218K - $335K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Familiarity with Kubernetes, Docker, Istio, Terraform, Prometheus, Grafana, TSDBs and observability pipelines (e.g. either for logging or metrics or tracing) * Skilled in defining and instrumenting ...

Engineering Manager Observability

Sunnyvale, CA · On-site

$218K - $335K/yr

  • Medical

  • Dental

  • Vision

  • Life

  • Retirement

  • PTO

Familiarity with Kubernetes, Docker, Istio, Terraform, Prometheus, Grafana, TSDBs and observability pipelines (e.g. either for logging or metrics or tracing) * Skilled in defining and instrumenting ...

Kubernetes Engineer

Sunnyvale, CA · On-site

$67.75 - $90/hr

Role: Kubernetes Engineer Location: Sunnyvale/Austin - Hybrid Duration: 6+ months Key ... Implement CI/CD, GitOps, observability , and leverage Helm, Kustomize, Service Mesh, Ingress ...

next page

Showing results 1-20

Observability Kubernetes information

What job categories do people searching Observability Kubernetes jobs in California look for?

The top searched job categories for Observability Kubernetes jobs in California are:

What cities in California are hiring for Observability Kubernetes jobs?

Cities in California with the most Observability Kubernetes job openings:

Lead Site Reliability Engineer, Factory Software

Tesla

Fremont, CA • On-site

$62.50 - $83.25/hr

Full-time

Re-posted 12 days ago


Tesla rating

8.5

Company rating: 8.5 out of 10

Based on 681 frontline employees who took The Breakroom Quiz

1st of 44 rated automakers


Job description

Job Summary:
Tesla is building critical applications to enable manufacturing and warehouse management with a strong emphasis on reliability, availability, scalability, speed, and security. As the Lead Site Reliability Engineer, you will be the primary technical owner and leader for the Factory Software team’s reliability, observability, and infrastructure strategy, combining deep hands-on engineering with leadership to ensure the full stack is highly reliable and performant.
Responsibilities:
• Provide technical leadership and set the vision for observability, reliability, and platform standardization across the Factory Software team
• Design and implement end-to-end observability and telemetry solutions (OTEL, Prometheus, Grafana, Tempo, etc.) while mentoring the team on best practices
• Own the reliability of the full stack: Kubernetes infrastructure, virtual machines, databases, and the middleware applications connecting PLCs, MES systems, and other factory services
• Define and drive SLIs, SLOs, error budgets, and golden signals across services
• Lead major initiatives to eliminate speed bottlenecks, database contention, and infrastructure issues through proactive monitoring and automation
• Write production-grade code and build tools to reduce toil and improve deployment, monitoring, and operational workflows
• Participate hands-on in on-call rotations, live troubleshooting during outages (NOC bridges), and blameless post-mortems
• Collaborate closely with Platform Engineering, Infrastructure, Controls Engineering, and Software Engineering teams to embed reliability and observability into architecture and development practices
• Mentor and coach engineers on technical excellence, observability, Kubernetes, Linux, networking, and reliable system design
• Drive continuous improvement in incident response, system performance, and engineering standards across the team
Qualifications:
Required:
• 7+ years of experience in Site Reliability Engineering, Platform Engineering, or related systems roles, with significant hands-on experience at scale
• Strong technical expertise in Kubernetes, Docker, Linux administration, and networking (routing, VLANs, firewalls, load balancers)
• Deep experience with observability tools and concepts (Prometheus, Grafana, Tempo, OTEL, Splunk, etc.)
• Proven track record of designing and implementing reliable, observable distributed systems
• Proficiency in at least one high-level language (Go, Python, or Java) with experience writing production-grade code
• Demonstrated ability to lead technical initiatives and raise the engineering bar without formal people management authority
• Experience with on-call rotations, incident command, and driving reliability improvements through blameless post-mortems
• Strong bias for action, excellent communication skills, and a desire to mentor and uplift other engineers
Preferred:
• Experience in manufacturing, industrial automation, or complex operational environments is a strong plus
Company:
Tesla is an electric vehicle and clean energy company that provides electric cars, solar, and renewable energy solutions. Founded in 2003, the company is headquartered in Austin, USA, with a team of 10001+ employees. The company is currently Late Stage.

What Tesla employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom