1

Senior Ai Infrastructure Engineer Jobs (NOW HIRING)

AI Infrastructure Engineer

San Jose, CA · On-site

$126K - $165K/yr

About the Position We are looking for a senior AI Inference Infrastructure Software Engineer with strong hands-on experience building, optimizing, and deploying high-performance, scalable inference ...

AI Infrastructure Engineer

San Jose, CA · On-site

$126K - $165K/yr

About the Position We are looking for a senior AI Inference Infrastructure Software Engineer with strong hands-on experience building, optimizing, and deploying high-performance, scalable inference ...

Deploy and manage AI services across cloud and on-prem environments. * Automate infrastructure ... Collaborate with engineering, data, and product teams to support AI initiatives. * Excellent ...

AI Infrastructure Engineer

Charlotte, NC · On-site

$105K - $137K/yr

Deploy and manage AI services across cloud and on-prem environments. * Automate infrastructure ... Collaborate with engineering, data, and product teams to support AI initiatives. * Excellent ...

AI Infrastructure Engineer

Fremont, CA · On-site

$126K - $165K/yr

Development and deployment of AI infrastructure workloads on-prem (GPU scheduling, model serving ... infra engineering Benefits * Medical Insurance * Dental Insurance * Vision Insurance * 401(k)

AI Infrastructure Engineer

Fremont, CA · On-site

$126K - $165K/yr

Development and deployment of AI infrastructure workloads on-prem (GPU scheduling, model serving ... infra engineering Benefits * Medical Insurance * Dental Insurance * Vision Insurance * 401(k)

AI Infrastructure Engineer

Fremont, CA

$117K - $154K/yr

Development and deployment of AI infrastructure workloads on-prem (GPU scheduling, model serving ... infra engineering Benefits * Medical Insurance * Dental Insurance * Vision Insurance * 401(k)

AI Infrastructure Engineer

San Francisco, CA · On-site

$126K - $166K/yr

Spellbrush, the world's leading generative AI studio behind niji・journey , is looking for an AI Infrastructure Engineer to join us in building out end-to-end ML infrastructure to run our models on ...

Showing results 41-60

Senior Ai Infrastructure Engineer information

See salary details

$22.5K

$127K

$175.5K

How much do senior ai infrastructure engineer jobs pay per year?

As of Aug 14, 2026, the average yearly pay for senior ai infrastructure engineer in the United States is $126,969.00, according to ZipRecruiter salary data. Most workers in this role earn between $108,500.00 and $147,500.00 per year, depending on experience, location, and employer.

What does a senior AI infrastructure engineer do?

A Senior AI Infrastructure Engineer is responsible for designing, building, and maintaining the large-scale computing systems that support artificial intelligence (AI) and machine learning (ML) workloads. They work on optimizing data pipelines, managing cloud or on-premise infrastructure, ensuring scalability, and enabling efficient training and deployment of AI models. These professionals collaborate closely with data scientists, software engineers, and IT teams to create robust, high-performance environments that support the rapid development and deployment of AI solutions.

What are some typical challenges faced by senior AI infrastructure engineers when scaling AI systems for production?

Senior AI Infrastructure Engineers often encounter challenges related to managing large-scale data pipelines, ensuring low-latency model serving, and maintaining system reliability as user demand grows. Balancing resource allocation for compute-intensive workloads, optimizing infrastructure costs, and implementing robust monitoring are common hurdles. Collaboration with data scientists, DevOps, and product teams is crucial to streamline deployment cycles and rapidly address issues as they arise. Mastery of distributed systems and cloud platforms often distinguishes top performers in this role.

What are the key skills and qualifications needed to thrive as a senior AI infrastructure engineer?

To thrive as a Senior AI Infrastructure Engineer, you need deep expertise in computer science, cloud computing, distributed systems, and AI/ML frameworks, often supported by a relevant degree and significant experience. Proficiency with tools such as Kubernetes, Docker, TensorFlow, PyTorch, and cloud platforms like AWS or Azure—as well as experience with CI/CD pipelines—is typically required. Strong problem-solving abilities, collaboration, and effective communication are standout soft skills for this role. These competencies are crucial for designing scalable, reliable AI infrastructure that supports complex machine learning workflows and organizational goals.
More about Senior Ai Infrastructure Engineer jobs

What cities are hiring for Senior Ai Infrastructure Engineer jobs?

Cities with the most Senior Ai Infrastructure Engineer job openings:

What are the most commonly searched types of Ai Infrastructure Engineer jobs?

The most popular types of Ai Infrastructure Engineer jobs are:

What states have the most Senior Ai Infrastructure Engineer jobs?

States with the most job openings for Senior Ai Infrastructure Engineer jobs include:

Infographic showing various Senior Ai Infrastructure Engineer job openings in the United States as of August 2026, with employment types broken down into 100% Full Time. Highlights an 60% In-person, and 40% Remote job distribution, with an average salary of $126,969 per year, or $61 per hour.

AI Infrastructure Engineer

NIO

San Jose, CA • On-site

$126K - $165K/yr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Re-posted 21 days ago


Job description

JOB DESCRIPTION
About NIO
NIO is a pioneer and a leading company in the premium smart electric vehicle market. Founded in November 2014, NIO's mission is to shape a joyful lifestyle. NIO aims to build a community starting with smart electric vehicles to share joy and grow together with users.
NIO designs, develops, jointly manufactures and sells premium smart electric vehicles, driving innovations in next-generation technologies in autonomous driving, digital technologies, electric powertrains and batteries. NIO differentiates itself through its continuous technological breakthroughs and innovations, such as its industry-leading battery swapping technologies, Battery as a Service, or BaaS, as well as its proprietary autonomous driving technologies and Autonomous Driving as a Service, or ADaaS.
NIO's product portfolio consists of the ES8, a six-seater smart electric flagship SUV, the ES7 (or the EL7), a mid-large five-seater smart electric SUV, the ES6, a five-seater all-round smart electric SUV, the EC7, a five-seater smart electric flagship coupe SUV, the EC6, a five-seater smart electric coupe SUV, the ET7, a smart electric flagship sedan, and the ET5, a mid-size smart electric sedan.
About the Position
We are looking for a senior AI Inference Infrastructure Software Engineer with strong hands-on experience building, optimizing, and deploying high-performance, scalable inference systems. This position is focused on designing, implementing, and delivering production-grade software that powers real-world applications of Large Language Models (LLMs) and Vision-Language Models (VLMs).
This is an exciting opportunity for an engineer who thrives at the intersection of AI systems, hardware acceleration, and large-scale robust deployment, and who wants to see their contributions ship in production, at scale.
In this role, you will directly shape the architecture, roadmap and performance of AI capabilities of our AIOS platform, driving innovations that make LLM/VLM systems fast, efficient, and scalable across cloud, edge, and hybrid edge-cloud environments. You will work closely with system, hardware, and product teams to deliver high-performance inference kernels for hardware accelerators, design scalable inference serving systems, and integrate optimizations such tensor parallelism and custom kernels into production pipelines. Your work will have immediate impact, powering intelligent automotive systems in the next generation of electric vehicles.
Roles and Responsibilities:
  • Design and implement high-performance, scalable inference systems for LLMs and VLMs across cloud, edge, and edge-cloud hybrid platforms.
  • Develop and optimize custom kernels and operators for specific hardware accelerators (GPU, NPU, DSP, etc.), improving throughput, latency, and memory efficiency.
  • Integrate advanced optimization techniques such as KV-cache management, tensor/model parallelism, quantization, and memory-efficient execution into production inference systems.
  • Partner with system and hardware teams to ensure tight hardware-software integration and optimal performance across diverse compute environments.
  • Translate architectural requirements into robust, maintainable, production-ready software that meets performance, safety, and reliability standards.
  • Define and drive the evolution roadmap for LLM/VLM inference in the AIOS stack, ensuring scalability and adaptability to new workloads.
  • Stay ahead of industry trends and competitor solutions, applying best practices from both AI and large-scale systems engineering.

Must Qualifications:
  • 5+ years of hands-on software development experience in building and optimizing AI inference systems at scale.
  • Direct experience in LLM/VLM model internals, including Transformer-based architectures, inference bottlenecks, and optimization techniques.
  • Strong expertise in performance engineering: kernel development, parallelism strategies, memory optimization, and distributed inference systems.
  • Proficiency with GPU/NPU programming (CUDA, or vendor-specific SDKs), compiler toolchains, and deep learning frameworks (PyTorch, or TensorFlow).
  • Strong programming skills in C/C++, with a track record of delivering high-performance, production-grade software.
  • Solid foundation in computer architecture, systems programming (CPU/GPU pipelines, memory hierarchy, scheduling), and embedded systems.
  • BS/MS in Computer Science, Computer Engineering, or related technical field.
  • Excellent communication and collaboration skills, with the ability to work across cross-functional teams.

Preferred Qualifications:
  • Master's or PhD degree in Computer Science, Electrical/Computer Engineering, or related fields, plus 5 years industry experience
  • Experience building inference serving systems for large models, including batching, scheduling, caching, and load balancing.
  • Expertise in hardware-aware model optimization (e.g., kernel fusion, mixed precision, quantization, pruning).
  • Familiarity with edge and embedded AI, including real-time constraints and limited-resource optimization.
  • Contributions to widely used AI frameworks, libraries, or performance-critical software (open source or proprietary).

Compensation:
The US base salary range for this full-time position is $192,100.00 - $249,600.00.
  • Within the range, individual pay is determined by work location and additional factors, including job-related skills, experience, and relevant education or training.
  • Please note that the compensation details listed in US role postings reflect the base salary only. It does not include discretionary bonus, equity, or benefits.

Benefits:
Along with competitive pay, as a full-time NIO employee, you are eligible for the following benefits on the first day you join NIO:
  • Anthem Blue Cross, HSA, and Kaiser HMO medical plans with $0 for Employee Only Coverage.
  • Dental (including orthodontic coverage) and vision plan. Both provide options with a $0 paycheck contribution covering you and your eligible dependents.
  • Company Paid HSA (Health Savings Account) Contribution when enrolled in the High Deductible Anthem Blue Cross medical plan
  • Healthcare and Dependent Care Flexible Spending Accounts (FSA)
  • 401(k) with Brokerage Link option
  • Company paid Basic Life, AD&D, short-term and long-term disability insurance
  • Employee Assistance Program
  • Sick and Vacation time
  • 13 Paid Holidays a year
  • Paid Parental Leave for first 8 weeks at full pay (eligible after 90 days of employment with NIO)
  • Paid Disability Leave for first 6 weeks at full pay (eligible after 90 days of employment with NIO)
  • Voluntary benefits including: Voluntary Life and AD&D options for you, your spouse/domestic partner and dependent child(ren), pet insurance
  • Commuter benefits
  • Mobile Cell Phone Credit
  • Free lunch and snacks
  • Onsite gym
  • Employee discounts and perks program