Role: Senior SRE Platform Engineer with Kubernetes
Location: San Jose, CA (Hybrid)
Duration: Long Term
Job Summary
We are looking for an experienced Senior SRE Platform Engineer to lead the design, automation, and operational excellence of Kubernetes-based cloud platforms. The ideal candidate will drive platform reliability, scalability, and infrastructure automation across enterprise environments.
Key Responsibilities
- Design, build, and maintain highly available Kubernetes platforms.
- Lead infrastructure automation and CI/CD implementation.
- Improve platform reliability, performance, and disaster recovery capabilities.
- Drive observability initiatives using monitoring and logging solutions.
- Troubleshoot complex production issues and lead root cause analysis.
- Mentor junior engineers and promote SRE best practices.
- Collaborate with architecture, development, and security teams.
Required Skills
- 7 10 years of SRE, DevOps, or Platform Engineering experience.
- Advanced expertise in Kubernetes, Docker, and container platforms.
- Strong cloud experience with AWS, Azure, or Google Cloud Platform.
- Hands-on experience with Terraform, Helm, ArgoCD, or GitOps.
- Strong experience with CI/CD pipelines and Infrastructure as Code.
- Expertise in Prometheus, Grafana, ELK, Datadog, or Splunk.
- Strong scripting/programming skills in Python, Go, or Bash.