1

Stack Infrastructure Jobs in California (NOW HIRING)

Infrastructure Software Engineer

San Jose, CA ยท On-site

$150K - $250K/yr

We build this infrastructure as software - and we engineer it with the same best practices we apply ... You'll also architect and implement a state-of-the-art observability stack with LLM integration and ...

AI Infrastructure Engineer

Fremont, CA ยท On-site

$126K - $165K/yr

Build and maintain Infrastructure-as-Code using Terraform (Terragrunt), Ansible, etc ... Deploy, secure, and operate our HashiCorp stack (Vault, Boundary) alongside identity/access via ...

AI Infrastructure Engineer

Fremont, CA ยท On-site

$126K - $165K/yr

Build and maintain Infrastructure-as-Code using Terraform (Terragrunt), Ansible, etc ... Deploy, secure, and operate our HashiCorp stack (Vault, Boundary) alongside identity/access via ...

AI Infrastructure Engineer

Fremont, CA ยท On-site

$117K - $154K/yr

Build and maintain Infrastructure-as-Code using Terraform (Terragrunt), Ansible, etc ... Deploy, secure, and operate our HashiCorp stack (Vault, Boundary) alongside identity/access via ...

Data Infrastructure Engineer

San Francisco, CA ยท On-site

$126K - $166K/yr

Ability to work across the stack, from analyzing data to designing the systems around it * Willingness to travel to customer sites Preferred * Data infrastructure in a robotics, autonomy, or ...

Working across the full stack to solve complex problems independently * Mentoring junior engineers ... GPU infrastructure, and/or deep learning applications * 2+ years experience with ML platforms ...

Data Infrastructure Engineer

San Francisco, CA ยท On-site

$126K - $166K/yr

Ability to work across the stack, from analyzing data to designing the systems around it * Willingness to travel to customer sites Preferred * Data infrastructure in a robotics, autonomy, or ...

Position Overview We are seeking a Full Stack Developer to help build and maintain our cloud infrastructure and web applications. This role requires expertise in modern web technologies and a strong ...

Infrastructure Engineer

San Francisco, CA ยท On-site +1

$100K - $250K/yr

AI Agent , the AI harness that enables anyone to build powerful full-stack applications * App ... Infrastructure : scale large, data-intensive, multi-tenant, distributed services * Product sense ...

Infrastructure Engineer

San Francisco, CA ยท On-site +1

$126K - $166K/yr

Experience working in a high growth, modern web tech stack NICE-TO-HAVES * Experience migrating ... Experience Forecasting infrastructure requirements ROLE DESCRIPTION As we grow our user base our ...

Infrastructure Engineer

San Francisco, CA ยท On-site

$100K - $250K/yr

AI Agent , the AI harness that enables anyone to build powerful full-stack applications * App ... Infrastructure : scale large, data-intensive, multi-tenant, distributed services * Product sense ...

Software Engineer, Infrastructure

San Francisco, CA ยท On-site

$203K - $241K/yr

Work across layers of the stack--debugging system bottlenecks, evolving core infrastructure, and solving novel problems in performance and scalability. * Reliability Engineering: Build scalable ...

About the Role The Streaming team owns the full-stack infrastructure powering the video experience for over 1.4M Verkada cameras. We take on complex streaming challenges to deliver live and recorded ...

Showing results 41-60

Stack Infrastructure information

See California salary details

$43.9K

$133K

$188K

How much do stack infrastructure jobs pay per year?

As of Sep 13, 2026, the average yearly pay for stack infrastructure in California is $133,006.00, according to ZipRecruiter salary data. Most workers in this role earn between $109,500.00 and $155,900.00 per year, depending on experience, location, and employer.

What is stack infrastructure?

Stack Infrastructure is a company that specializes in providing data center solutions, including colocation and cloud infrastructure services. They design, build, and operate data centers for businesses requiring secure, reliable, and scalable IT environments. Stack Infrastructure supports enterprises with high-performance computing needs and offers customizable infrastructure to accommodate growth and technological advancements.

What are the key skills and qualifications needed to thrive in stack infrastructure roles?

To excel in Stack Infrastructure roles, you need a solid understanding of data center operations, networking, hardware management, and often a degree in IT or related fields. Familiarity with tools like virtualization platforms (VMware, Hyper-V), monitoring systems, and certifications such as CompTIA Server+, Cisco CCNA, or AWS Certified Solutions Architect are typically required. Strong problem-solving abilities, attention to detail, and effective communication make candidates stand out in this position. These skills ensure the reliable and secure operation of critical infrastructure, supporting business continuity and scalable technology solutions.

What are some common challenges faced by professionals working in stack infrastructure roles, and how can they be addressed?

Professionals in Stack Infrastructure often face challenges such as managing scalability, ensuring system reliability, and coordinating between development and operations teams. Adapting to rapidly evolving technologies and maintaining security standards can also be demanding. Overcoming these challenges typically involves staying current with industry best practices, utilizing automation tools, and fostering clear communication within cross-functional teams. Continuous learning and proactive problem-solving are key to thriving in this dynamic environment.

Is Stack Infrastructure a good company to work for?

Stack Infrastructure offers roles related to infrastructure management, often involving skills in cloud platforms, networking, and system administration. Employee reviews indicate a focus on technical growth and collaborative environments, but experiences can vary based on role and location.

What cities in California are hiring for Stack Infrastructure jobs?

Cities in California with the most Stack Infrastructure job openings:

Infographic showing various Stack Infrastructure job openings in California as of September 2026, with employment types broken down into 1% Internship, 85% Full Time, 12% Part Time, and 2% Contract. Highlights an 84% Physical, 4% Hybrid, and 12% Remote job distribution, with an average salary of $133,006 per year, or $63.9 per hour.

Infrastructure Software Engineer

San Jose, CA โ€ข On-site

$150K - $250K/yr

Full-time

Medical, Dental, Vision

Re-posted 13 hours ago


Job description

About Etched

Etched is building hardware for frontier intelligence. We co-design chips, racks, software, and manufacturing to deliver best-in-class throughput and latency across both prefill and decode workloads. Our first products are heavily focused on inference. Backed by hundreds of millions from top-tier investors and staffed by leading engineers, Etched is redefining the infrastructure layer for the fastest growing industry in history.

Job Summary

Building cutting-edge model-specific ASICs requires crafting custom infrastructure and toolchains to support ultra-fast, reliable, and scalable development across the stack - from simulation to silicon. We build this infrastructure as software - and we engineer it with the same best practices we apply to our products. We use the same rigor, design discipline, and quality standards and testing as we do to our ASIC, software, and platform.

You will lead the development and adoption of next-generation infrastructure tooling, enabling Etched ASIC, Software, and Platform engineers to iterate faster, build more reliably, and push the boundaries of AI performance. This includes building and scaling our hybrid high-performance compute (HPC) cluster, optimized for massively parallel CI, EDA workflows, Emulation, and hardware-aware job execution.

You’ll also architect and implement a state-of-the-art observability stack with LLM integration and a strong emphasis on streaming health and performance telemetry, log aggregation, distributed tracing, insight generation, synthetic testing, and smart alerting - across CI pipelines, simulation clusters, and service endpoints.

This role demands a strong software engineering mindset, quality instincts, and deep understanding of systems. It’s not just about writing scripts - it’s about writing code that builds and manages infrastructure with precision, repeatability, and intent.

Key responsibilities

  • Design and build the orchestration layers that drive our hybrid high-performance clusters—enabling simulation, synthesis, and continuous integration of AI ASICs at unprecedented scale.

  • Develop and maintain a fully programmable infrastructure control plane to ensure reproducibility, auditability, and rapid iteration across the entire stack.

  • Create tools and abstractions that empower engineers to harness massive parallelism without worrying about the underlying complexity..

  • Prototype and execute workload orchestration and migration strategies between on-premise and cloud environments, balancing performance, storage availability and replication, uptime, and cost across heterogeneous hardware and compute backends.

  • Implement real-time telemetry, tracing systems that surface insights from millions of metrics, enabling proactive debugging and system optimization.

  • Build a full observability stack that includes dashboards, alerting, automated responses, and a synthetic testing framework to proactively test infrastructure performance and reliability for various application and data flows, ensuring we remain proactive against issues impacting development and productivity workflows.

Representative projects

  • Design and deploy a fully automated, scalable hybrid HPC cluster, combining bare-metal servers and switches with cloud instances, provisioned through MaaS and orchestrated via SLURM and Kubernetes, optimized for mixed EDA workloads and parallel CI pipelines.

  • Develop a real-time observability system for ASIC toolchain jobs and distributed builds, integrating Prometheus, Grafana, and VictoriaMetrics with streaming telemetry, tracing, and alerting to detect performance regressions before they hit silicon.

  • Architect and implement a programmable infrastructure-as-code control plane, using Terraform, Ansible, and Puppet, to version, audit, and redeploy every layer of Etched's development stack with deterministic reproducibility.

  • Create a zero-downtime interactive development environment that provisions and connects Jupyter and VS Code sessions to GPUs and high-memory nodes via a secure zero-trust network, abstracting away cluster state and machine failures.

  • Prototype and evaluate dynamic workload migration strategies between on-premise and cloud environments to optimize for latency, reliability, and cost across simulation and synthesis pipelines.

  • Design a synthetic testing and fault injection framework to validate the behavior of infrastructure under high-load, degraded hardware, and intermittent network partitions - before they happen in production.

You may be a good fit if you

  • Are a systems-minded software engineer who loves building foundational platforms, working close to the metal and cloud, solving high-leverage problems at scale.

  • Are a deeply technical engineer who treats infrastructure as a software problem - prioritizing clean abstractions, version control,small change lists, easy roll backs, testing, and long-term maintainability over ad hoc configuration.

  • Have strong programming skills in languages such as Python, Go, Rust, and C++, and are comfortable building production-grade tooling.

  • Possess expert-level knowledge of Linux, virtualization, containerization, and CI/CD pipelines, with a deep understanding of how to debug, optimize, and scale complex systems.

  • Are familiar with Infrastructure as Code tools like OpenTofu, Ansible, or Puppet, and enjoy designing declarative, reproducible infrastructure systems.

  • Understand and use PromQL and other telemetry/query languages and have used LLM to extract insight from real-time metrics, and know how to architect and tune observability stacks.

  • Have a track record of debugging and resolving difficult hardware-software integration problems across bare-metal systems, networks, and distributed workloads.

  • Can lead and mentor technical teams, guiding design decisions and helping others develop sound engineering instincts.

  • Have 8+ years of experience in infrastructure engineering, systems programming, or backend software development - ideally in environments where performance, scale, or hardware interaction mattered.

  • Are driven by curiosity, take initiative, and have an innate sense of ownership — you thrive in uncharted territory, design for edge cases, and love making systems more powerful, reliable, and elegant.

Strong candidates may also have experience with

  • Familiarity with Bazel build system

  • Deep understanding of ASIC development flows, especially those involving Synopsys, Cadence, and Verilator, including how EDA tools interact with infrastructure for simulation, synthesis, and verification.

  • Hands-on experience architecting systems with AWS, GCP, or Azure, including hybrid on-prem/cloud deployments, workload migration strategies, and cloud-native orchestration tooling.

  • Experience monitoring, provisioning, and debugging bare-metal servers, network hardware, and high-performance storage systems in rack-scale environments.

  • Comfortable in profiling and optimizing compute environments for single-threaded latency, memory-bound workloads, or I/O throughput, especially in the context of simulation or CI performance.

  • Proficiency building or operating telemetry systems at scale using Prometheus, Grafana, Loki, VictoriaMetrics, and tools for distributed tracing, log aggregation, and real-time alerting across heterogeneous mediums (SMS, email, push alerts, etc.)

Benefits

  • Medical, dental, and vision packages with generous premium coverage

    • $500 per month credit for waiving medical benefits

  • Housing subsidy of $2k per month for those living within walking distance of the office

  • Relocation support for those moving to San Jose (Santana Row)

  • Various wellness benefits covering fitness, mental health, and more

  • Daily lunch + dinner in our office

  • Unlimited compute budget subject to ROI justification

How we’re different

Etched believes in the Bitter Lesson. We are the first inference-focused frontier AI system. Our addressable market is the entirety of inference, unlike many of our competitors.

 

We are a fully in-person team in San Jose (Santana Row), and greatly value engineering skills. We do not have boundaries between engineering and research, and we expect all of our technical staff to contribute to both and work across disciplines as needed.

Compensation Range: $150K - $250K