1

Gpu Infrastructure Engineer Jobs (NOW HIRING)

The role Nebius is looking for a Sr. Solutions Engineer, you will act as the primary technical partner for customers deploying and operating GPU clusters and AI infrastructure on Nebius. You will ...

AI Infrastructure Engineer L3

Santa Clara, CA ยท On-site

$125K - $164K/yr

AI Infrastructure Engineer L3 Location: Santa Clara, CA Experience: 10 plus Years Employment Type ... Deploy and manage NVIDIA GPU infrastructure (A100, H100, L40) and AI accelerator platforms.

The role Nebius is looking for a Sr. Solutions Engineer, you will act as the primary technical partner for customers deploying and operating GPU clusters and AI infrastructure on Nebius. You will ...

AI Infrastructure Engineer

New York, NY ยท Remote

$140K - $165K/yr

As an AI Infrastructure Engineer, your role will include: * Lead Technical Deployments: Drive end ... Configure and troubleshoot bare metal GPU node infrastructure, including CNI configuration, GPU ...

Designing, building, and scaling GPU compute infrastructure, training frameworks, and ... Mentoring junior engineers and contributing to the growth of the team BASIC QUALIFICATIONS:

next page

Showing results 1-20

Gpu Infrastructure Engineer information

See salary details

$46.5K

$127.1K

$182K

How much do gpu infrastructure engineer jobs pay per year?

As of Sep 11, 2026, the average yearly pay for gpu infrastructure engineer in the United States is $127,066.00, according to ZipRecruiter salary data. Most workers in this role earn between $107,500.00 and $141,000.00 per year, depending on experience, location, and employer.

What are popular job titles related to Gpu Infrastructure Engineer jobs?

For Gpu Infrastructure Engineer jobs, the most frequently searched job titles are:

Infographic showing various Gpu Infrastructure Engineer job openings in the United States as of August 2026, with employment types broken down into 94% Full Time, 2% Part Time, and 4% Contract. Highlights an 85% Physical, 5% Hybrid, and 10% Remote job distribution, with an average salary of $127,066 per year, or $61.1 per hour.

Senior GPU Infrastructure Engineer

San Francisco, CA โ€ข On-site

$127K - $173K/yr

Full-time

Re-posted 19 days ago


Job description

Job Summary:
Hyperbolic Labs is on a mission to democratize AI by breaking down the barriers to computing power with their Open-Access AI Cloud. They are seeking a Senior Infrastructure Engineer to help build and scale their GPU Cloud Marketplace, transforming raw GPUs into a programmable pool for AI developers and researchers.
Responsibilities:
โ€ข help build and scale Hyperbolic's GPU Cloud Marketplace
โ€ข building a multi-tenancy provisioning and virtualization solution
โ€ข transforming raw GPUs from diverse global suppliers into a programmable, orchestrated pool that serves thousands of AI developers and researchers
โ€ข work at the cutting edge of cloud infrastructure
โ€ข building the core orchestration layer that enables our platform to deliver up to 75% cost savings compared to traditional cloud providers
Qualifications:
Required:
โ€ข Deep understanding of bare-metal provisioning and lifecycle management, including IPMI/Redfish, BMC-based remote management, PXE boot, and automated OS deployment workflows
โ€ข Deep understanding of GPU scheduling and orchestration, including GPU type awareness, memory management, topology considerations, placement strategies for multi-GPU jobs, and fragmentation minimization
โ€ข Strong infrastructure and DevOps engineering skills with proficiency in Terraform or Pulumi, CI/CD for infrastructure, secrets management, configuration management, and observability stack implementation
โ€ข Experience with storage and data infrastructure for AI/ML workloads, including object storage, high-IOPS block storage, and distributed file systems for training data and checkpoints
โ€ข Proficiency with API design and cloud-init for automated provisioning and configuration
โ€ข Solid understanding of GPU architecture, CUDA, and GPU compute optimization
โ€ข Highly collaborative team player with excellent communication skills across technical and non-technical stakeholders
โ€ข Proven ability to work effectively with hardware vendors and vendor engineering teams to troubleshoot issues and optimize integrations
โ€ข Experience building and scaling cloud infrastructure or distributed systems in production environments
Preferred:
โ€ข Familiarity with high-performance networking technologies such as InfiniBand and RoCE (RDMA over Converged Ethernet)
โ€ข Experience with distributed storage systems such as Ceph, Weka, or VAST Data
Company:
Hyperbolic is the open-access AI cloud made for AI developers, providing fast, affordable access to compute, inference, and AI services. Founded in 2022, the company is headquartered in Irvine, USA, with a team of 11-50 employees. The company is currently Early Stage.