Role We are looking for a Staff Software Engineer (Service Platform & Orchestration) to join our ... and performance optimization * Proven track record of developing Platform APIs (REST/gRPC) with ...

31 Zscaler Performance Engineer Jobs Hiring Near You
Role We are looking for a Staff Software Engineer (Service Platform & Orchestration) to join our ... and performance optimization * Proven track record of developing Platform APIs (REST/gRPC) with ...
Senior Staff Software Engineer (C / Networking / Dataplane)
San Jose, CA · Hybrid
$143K - $189K/yr
The Senior Staff Software Engineer (C / Networking / Dataplane) will design and implement high-performance data path features for load balancing and L7 protocols, directly enhancing the scalability ...
Senior Staff Software Engineer (C / Networking / Dataplane)
San Jose, CA · Hybrid
$143K - $189K/yr
The Senior Staff Software Engineer (C / Networking / Dataplane) will design and implement high-performance data path features for load balancing and L7 protocols, directly enhancing the scalability ...
Develop and maintain high-performance resilient UI using ReactJS, TypeScript, and Tailwind ... engineering excellence and driving innovation What Will Make You Stand Out (Preferred ...
Develop and maintain high-performance resilient UI using ReactJS, TypeScript, and Tailwind ... engineering excellence and driving innovation What Will Make You Stand Out (Preferred ...
Principal Rust Developer
San Jose, CA · Hybrid
Optimize system performance through profiling tools across kernel-space and user-space * Engage in code reviews, system design discussions, technical documentation, and mentoring junior engineers Who ...
Principal Rust Developer
San Jose, CA · Hybrid
Optimize system performance through profiling tools across kernel-space and user-space * Engage in code reviews, system design discussions, technical documentation, and mentoring junior engineers Who ...
Staff Site Reliability Engineer
San Jose, CA · Hybrid
$66.75 - $88.75/hr
You are an SRE with proven experience in Linux/UNIX System Administration and hands-on expertise in ... Analyze and troubleshoot systems performance and complex issues across operating systems and ...
Staff Site Reliability Engineer
San Jose, CA · Hybrid
$66.75 - $88.75/hr
You are an SRE with proven experience in Linux/UNIX System Administration and hands-on expertise in ... Analyze and troubleshoot systems performance and complex issues across operating systems and ...
Systematically enhance performance across the entire stack, including LLM models, by employing ... Deep experience in systems programming using Rust, with a focus on asynchronous frameworks such as ...
Systematically enhance performance across the entire stack, including LLM models, by employing ... Deep experience in systems programming using Rust, with a focus on asynchronous frameworks such as ...
Mentor and provide technical guidance to junior and mid-level engineers, promoting best practices and ensuring code quality * Continuously improve and optimize the performance, accessibility, and ...
Mentor and provide technical guidance to junior and mid-level engineers, promoting best practices and ensuring code quality * Continuously improve and optimize the performance, accessibility, and ...
... engineering, architecture, operations, and go-to-market teams to define and deliver infrastructure investments that strengthen platform performance, simplify operations, and support the needs of both ...
... engineering, architecture, operations, and go-to-market teams to define and deliver infrastructure investments that strengthen platform performance, simplify operations, and support the needs of both ...
Act as a critical thought partner to tech executives across AI, Product, Engineering, Innovation ... Drive executive-level succession planning, leadership coaching and development, and performance ...
Act as a critical thought partner to tech executives across AI, Product, Engineering, Innovation ... Drive executive-level succession planning, leadership coaching and development, and performance ...
Performance Metrics & Pipeline Management: Lead the sales operating cadence by driving pipeline ... Mentor a high-performing operations team and foster collaboration across product, engineering ...
Performance Metrics & Pipeline Management: Lead the sales operating cadence by driving pipeline ... Mentor a high-performing operations team and foster collaboration across product, engineering ...
... Reliability Engineering. The Procurement Operations Analyst executes end-to-end procurement ... This role improves operational efficiency and service-level performance through accurate tracking ...
... Reliability Engineering. The Procurement Operations Analyst executes end-to-end procurement ... This role improves operational efficiency and service-level performance through accurate tracking ...
Zscaler Jobs Information
What is it like to work at Zscaler?
What makes Zscaler an attractive place to work?
What other companies are hiring for Performance Engineer jobs?
What are the most popular jobs at Zscaler?
What are the most popular categories at Zscaler?

Full-time
Re-posted 6 days ago
Job description
Role
We are looking for a Staff Software Engineer (Service Platform & Orchestration) to join our team. This is a Hybrid (3 days in office) role, reporting to the VP, Engineering in the Zero Trust Exchange department. In this high-ownership position, you will engineer the management plane and reliability systems that govern ZIA's fleet lifecycle at massive scale.
This is a hands-on building role where you will lead the transformation of our infrastructure from legacy automation into a stateful, durable management plane (built on Temporal) to achieve deterministic "one-touch" provisioning and lifecycle operations. You will treat "infrastructure as a distributed system," developing self-healing capabilities and AI-driven SRE practices for a global fleet of 100k+ instances.
What you'll do (Role Expectations)
- Lead the hands-on development and migration to a workflow-as-code platform (Temporal), building replay-safe, idempotent workflows that ensure deterministic operations across a global scale
- Move the organization beyond "scripted automation" toward a robust management plane that treats the entire global fleet as a single, eventually-consistent distributed system
- Design and implement services that leverage LLMs and ML for intelligent signal correlation, automated triage, and "self-correcting" fleet operations
- Develop framework-level services and internal APIs that ensure all new products are delivered "orchestration-ready" with reliability hooks built directly into the code
- Build deep telemetry (metrics, traces, and events) into the management plane so that every fleet-wide action is fully explainable, auditable, and replayable
Who You Are (Success Profile)
- You thrive in ambiguity. You're comfortable building the path as you walk it. You thrive in a dynamic environment, seeing ambiguity not as a hindrance, but as the raw material to build something meaningful.
- You act like an owner. Your passion for the mission fuels your bias for action. You operate with integrity because you genuinely care about the outcome. You adapt to what's needed, navigating seamlessly between high-level strategy and hands-on execution.
- You operate with urgency. You understand that in a high-growth environment, speed and quality are not mutually exclusive. You have a relentless focus on execution and a bias for action, delivering high-impact results quickly to win for the customer and the team.
- You think at scale. You connect your day-to-day work to the larger company mission and think globally. You build solutions, processes, and teams that are not just effective today but are built to last and support a high-growth, global organization.
- You are resilient and adaptable. You view change as an opportunity and setbacks as temporary. You maintain composure and focus in high-pressure situations, guiding yourself and your team through complexity with a steady, positive hand.
What We're Looking for (Minimum Qualifications)
- Foundational understanding of AI/ML technologies and experience leveraging, securing, or positioning AI-driven solutions to optimize outcomes within your functional domain
- BS/MS in Computer Science or a related technical field with 5+ years of experience in hyperscale systems, with a deep understanding of the unique failure modes and technical hurdles that only emerge at massive scale
- Mastery of backend systems languages (Go, Java, Python, or others) with a proven ability to set the bar for code quality, maintainability, and distributed system correctness
- Strong experience designing and operating complex distributed systems, with a focus on solving systemic challenges in concurrency, failure handling, and performance optimization
- Proven track record of developing Platform APIs (REST/gRPC) with strong guarantees for idempotency, verification, and safe rollout patterns
- U.S. citizenship due to the nature of the customers assigned to this role
What Will Make You Stand Out (Preferred Qualifications)
- Proficiency with AI code-assistance tools (e.g., Cursor, Windsurf) to accelerate legacy refactoring and system development
- Proficiency in PostgreSQL, or other relational stores used for high-scale, stateful management-plane services
- Direct experience building or operating systems with Temporal.io, Cadence, or similar workflow engines
#LI-YC2 #LI-Hybrid
About Zscaler
Sourced by ZipRecruiter
Industry
Network security
Company size
501 - 1,000 Employees
Headquarters location
San Jose, CA, US
Year founded
2007