We are seeking a senior system software engineer to lead technically and own key parts of our low ... Prior GPU, CUDA, or GEMM experience is helpful for the role; deep experience developing low-level ...
We are seeking a senior system software engineer to lead technically and own key parts of our low ... Prior GPU, CUDA, or GEMM experience is helpful for the role; deep experience developing low-level ...
Senior System Software Engineer - Data Center Compute Diagnostics
Durham, NC · On-site
$118K - $156K/yr
We are seeking a senior system software engineer to lead technically and own key parts of our low ... Prior GPU, CUDA, or GEMM experience is helpful for the role; deep experience developing low-level ...
Senior System Software Engineer - Data Center Compute Diagnostics
Durham, NC · On-site
$118K - $156K/yr
We are seeking a senior system software engineer to lead technically and own key parts of our low ... Prior GPU, CUDA, or GEMM experience is helpful for the role; deep experience developing low-level ...
Prior GPU, CUDA, or GEMM experience is helpful but not required; you will have the opportunity to learn GPU architecture and programming while working with experienced engineers. You will own well ...
Prior GPU, CUDA, or GEMM experience is helpful but not required; you will have the opportunity to learn GPU architecture and programming while working with experienced engineers. You will own well ...
Senior Software Engineer, CUTLASS Kernels
$118K - $156K/yr
Experience with CUDA, OpenCL, HIP, SYCL, Mojo, Pallas, Triton, Mosaic, Halide, or any general-purpose or domain-specific programming language targeting highly parallel accelerators. * Deep ...
Senior Software Engineer, CUTLASS Kernels
$118K - $156K/yr
Experience with CUDA, OpenCL, HIP, SYCL, Mojo, Pallas, Triton, Mosaic, Halide, or any general-purpose or domain-specific programming language targeting highly parallel accelerators. * Deep ...
System Software Engineer - Data Center GPU Compute Diagnostics
Durham, NC · On-site
$167K - $198K/yr
In this role you will partner with a senior engineer leading the team's CUDA kernel and GEMM diagnostics work, owning well-scoped pieces of the codebase end-to-end while ramping on GPU ...
System Software Engineer - Data Center GPU Compute Diagnostics
Durham, NC · On-site
$167K - $198K/yr
In this role you will partner with a senior engineer leading the team's CUDA kernel and GEMM diagnostics work, owning well-scoped pieces of the codebase end-to-end while ramping on GPU ...
Familiarity with CUDA programming and/or GPUs. * Experience with HPC or large-scale computing environments. Your base salary will be determined based on your location, experience, and the pay of ...
Familiarity with CUDA programming and/or GPUs. * Experience with HPC or large-scale computing environments. Your base salary will be determined based on your location, experience, and the pay of ...
Senior System Software Engineer - Data Center GPU Compute Diagnostics
Durham, NC · On-site
$118K - $156K/yr
We are seeking a senior system software engineer to work on next-generation Data Center GPU ... Responsible for crafting CUDA/C++ diagnostic workloads and software infrastructure required for new ...
Senior System Software Engineer - Data Center GPU Compute Diagnostics
Durham, NC · On-site
$118K - $156K/yr
We are seeking a senior system software engineer to work on next-generation Data Center GPU ... Responsible for crafting CUDA/C++ diagnostic workloads and software infrastructure required for new ...
Senior Software Architect - Deep Learning and HPC Communications
Durham, NC · On-site
$125K - $170K/yr
Preferred : • Expertise in related technology and passion for what you do. • Experience with CUDA programming and NVIDIA GPUs. • Knowledge of high-performance networks like InfiniBand, RoCE ...
Senior Software Architect - Deep Learning and HPC Communications
Durham, NC · On-site
$125K - $170K/yr
Preferred : • Expertise in related technology and passion for what you do. • Experience with CUDA programming and NVIDIA GPUs. • Knowledge of high-performance networks like InfiniBand, RoCE ...
Training, Distributed inference, MoE, Reinforcement Learning, kernel authoring (on CUDA, Triton, cuTe, etc). Experience with programming for compute & communication overlap in distributed runtimes ...
Training, Distributed inference, MoE, Reinforcement Learning, kernel authoring (on CUDA, Triton, cuTe, etc). Experience with programming for compute & communication overlap in distributed runtimes ...
Senior Software Engineer, CUTLASS Platform
$118K - $156K/yr
Collaborate with GPU architecture, CUDA, and NVVM/PTX compiler teams to provide feedback on programming models and to assess the performance of future GPU hardware features. What we need to see:
Senior Software Engineer, CUTLASS Platform
$118K - $156K/yr
Collaborate with GPU architecture, CUDA, and NVVM/PTX compiler teams to provide feedback on programming models and to assess the performance of future GPU hardware features. What we need to see:
Rapid prototyping and development with Python, C++, CUDA or related DSLs (Triton, cuTe) * Solid ... Experience with parallel programming on at least one communication runtime (NCCL, NVSHMEM, MPI)
Rapid prototyping and development with Python, C++, CUDA or related DSLs (Triton, cuTe) * Solid ... Experience with parallel programming on at least one communication runtime (NCCL, NVSHMEM, MPI)
A grasp of the CUDA programming model and experience employing GPU profiling tools like NVIDIA Nsight Systems/Compute to address PCIe bottlenecks and kernel stalls. * Extensive knowledge of profiling ...
A grasp of the CUDA programming model and experience employing GPU profiling tools like NVIDIA Nsight Systems/Compute to address PCIe bottlenecks and kernel stalls. * Extensive knowledge of profiling ...
Drive architecture and execution that fully uses NVIDIA GPUs, CUDA, and the accelerated-computing ... Partner and developer interface. Demonstrated ability to work directly with external partners ...
Drive architecture and execution that fully uses NVIDIA GPUs, CUDA, and the accelerated-computing ... Partner and developer interface. Demonstrated ability to work directly with external partners ...
DevOps Engineer
$49.25 - $67.50/hr
About the job The Applied AI & Modeling (AAIM) Division seeks a Software Engineer skilled in GPU ... Knowledge CUDA enabled systems and GPU scheduling * Technology Savvy- Leveraging one's practical ...
DevOps Engineer
$49.25 - $67.50/hr
About the job The Applied AI & Modeling (AAIM) Division seeks a Software Engineer skilled in GPU ... Knowledge CUDA enabled systems and GPU scheduling * Technology Savvy- Leveraging one's practical ...
DevOps Engineer
Cary, NC · On-site
$49.25 - $67.50/hr
About the job The Applied AI & Modeling (AAIM) Division seeks a Software Engineer skilled in GPU ... Knowledge CUDA enabled systems and GPU scheduling * Technology Savvy- Leveraging one's practical ...
DevOps Engineer
Cary, NC · On-site
$49.25 - $67.50/hr
About the job The Applied AI & Modeling (AAIM) Division seeks a Software Engineer skilled in GPU ... Knowledge CUDA enabled systems and GPU scheduling * Technology Savvy- Leveraging one's practical ...
Senior Software Architect - Deep Learning and HPC Communications (Durham)
Durham, NC · On-site
$123K - $167K/yr
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high-performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such PyTorch, TensorFlow, etc.
Senior Software Architect - Deep Learning and HPC Communications (Durham)
Durham, NC · On-site
$123K - $167K/yr
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high-performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such PyTorch, TensorFlow, etc.
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high-performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such PyTorch, TensorFlow, etc.
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high-performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such PyTorch, TensorFlow, etc.
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high-performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such PyTorch, TensorFlow, etc.
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high-performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such PyTorch, TensorFlow, etc.
Senior Developer Technology Engineer - AI
Durham, NC · Hybrid
$52.75 - $69.50/hr
A background that includes parallel programming, e.g., CUDA, OpenACC, OpenMP, MPI, pthreads, etc. * Hands on experience doing low-level performance optimizations. * In-depth expertise with CPU and ...
Senior Developer Technology Engineer - AI
Durham, NC · Hybrid
$52.75 - $69.50/hr
A background that includes parallel programming, e.g., CUDA, OpenACC, OpenMP, MPI, pthreads, etc. * Hands on experience doing low-level performance optimizations. * In-depth expertise with CPU and ...
Experience with parallel programming, ideally CUDA C/C++ * Excellent communication and organization skills, with a logical approach to problem solving, time management, and task prioritization skills ...
Experience with parallel programming, ideally CUDA C/C++ * Excellent communication and organization skills, with a logical approach to problem solving, time management, and task prioritization skills ...
Cuda Engineer information
See Raleigh, NC salary details
$35.5K - $44.4K
5% of jobs
$44.4K - $53.3K
5% of jobs
$53.3K - $62.3K
0% of jobs
$62.3K - $71.2K
1% of jobs
$71.2K - $80.1K
5% of jobs
$86.8K is the 25th percentile. Wages below this are outliers.
$80.1K - $89K
11% of jobs
$89K - $98K
14% of jobs
$98K - $106.9K
7% of jobs
The median wage is $107.6K / yr.
$106.9K - $115.8K
13% of jobs
$115.8K - $124.7K
9% of jobs
$126.2K is the 75th percentile. Wages above this are outliers.
$124.7K - $133.7K
30% of jobs
$35.5K
$104.3K
$133.7K
How much do cuda engineer jobs pay per year?
What is a CUDA engineer?
What is the difference between Cuda Engineer vs GPU Developer?
| Aspect | Cuda Engineer | GPU Developer |
|---|---|---|
| Required Credentials | Bachelor's or Master's in Computer Science, Engineering, or related; knowledge of CUDA, C++, parallel programming | Bachelor's or Master's in Computer Science, Engineering, or related; experience with GPU programming, CUDA, OpenCL |
| Work Environment | Research labs, tech companies, hardware firms focusing on GPU acceleration | Software development teams, gaming, AI, scientific computing sectors |
| Employer & Industry Usage | Hardware manufacturers, AI companies, high-performance computing firms | Game development, scientific research, machine learning applications |
While both roles involve GPU programming and CUDA expertise, a Cuda Engineer primarily focuses on developing and optimizing CUDA-based solutions for hardware acceleration. In contrast, a GPU Developer works on broader GPU programming tasks, including application development across various platforms. The roles often overlap but differ in scope and specific focus areas.
How much do Cuda engineers make?
What are some common challenges faced by CUDA engineers when optimizing GPU-accelerated applications?
What are the key skills and qualifications needed to thrive as a CUDA engineer?
Are CUDA engineers in demand?

$118K - $156K/yr
Full-time
Posted 5 days ago
Nvidia rating
9.6
Based on 17 frontline employees who took The Breakroom Quiz
8th of 242 rated software companies
Job description
We are seeking a senior system software engineer to lead technically and own key parts of our low-level diagnostic software. This software supports next-generation data center GPUs and rack-scale AI systems. Our team builds software that exercises and validates complex hardware, including processing units, storage and cache architectures, NICs, PCIe and NVLink interfaces, power delivery, and thermal behavior. This role is well suited to a senior embedded, firmware, device-driver, hardware-validation, or systems software engineer with a record of leading complex software projects that directly interface with hardware. Relevant experience may come from GPUs, CPUs, networking, storage, servers, embedded systems, or other complex silicon-based products. Prior GPU, CUDA, or GEMM experience is helpful for the role; deep experience developing low-level software for complex hardware systems is required.
This is a hands-on software development role in which you will architect, implement, debug, and maintain key components of the diagnostic software through validation, productization, and field support. In addition to making substantial individual code contributions, you will lead complex development efforts, mentor other engineers, and collaborate with hardware architects, driver developers, silicon-validation engineers, manufacturing teams, and field engineers to bring up new hardware and diagnose difficult system failures. Join an exciting, rewarding, and fast-moving environment building the systems that power modern AI.
What you'll be doing:
Architecting and developing diagnostic and stress software in C/C++ and Python for complex hardware systems.
Leading development efforts across multiple engineers, breaking ambiguous problems into actionable work, and mentoring engineers in low-level software development and debugging.
Interfacing with hardware blocks, firmware, Linux device drivers, registers, telemetry, and low-level debugging tools.
Assessing new hardware features and defining effective diagnostic and stress strategies for engineering validation, manufacturing, product qualification, and field use.
Designing targeted tests for compute engines, memory and cache subsystems, DMA engines, NICs, PCIe/NVLink interfaces, power, and thermal behavior.
Developing diagnostic and stress workloads ranging from low-level tests for GPU hardware to higher-level AI workloads using CUDA programming, GEMM-style compute, NCCL, and PyTorch.
Investigating complex hardware and software failures involving memory errors, ECC, data integrity, performance, thermals, voltage/frequency behavior, and high-speed interfaces.
Using modern development and analysis tools, including AI-assisted tools where appropriate, to accelerate coding, debugging, test creation, and failure analysis.
What we need to see:
BS or MS degree in Electrical Engineering, Computer Engineering, Computer Science, or a related field, or equivalent experience.
12+ years of experience in embedded software, firmware, Linux device drivers, systems software, hardware validation, diagnostics, or silicon bring-up.
Experience providing technical leadership for a complex software component or project, including coordinating work among various engineers and mentoring others.
Strong programming skills in C and C++, plus working proficiency in Python.
Extensive experience developing software that interacts with hardware, firmware, device drivers, hardware registers, or other low-level interfaces.
Background with PCIe, NVLink, or networking technologies such as Ethernet or InfiniBand.
Strong understanding of computer architecture concepts such as memory systems, caches, interrupts, DMA, buses, device I/O, bandwidth constraints, and hardware error behavior.
Experience debugging complex failures across hardware, firmware, device drivers, operating systems, and applications.
Ability to define technical direction, make sound engineering tradeoffs, and drive ambiguous problems through completion across organizational boundaries.
Excellent written and verbal communication skills, including the ability to communicate effectively with hardware architects, software engineers, manufacturing teams, field engineers, and technical leadership.
Widely considered to be one of the technology world's most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/
You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.About Nvidia
Sourced by ZipRecruiter
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.
Industry
Computer and electronic product manufacturing
Company size
10,000+ Employees
Headquarters location
Santa Clara, CA, US
Year founded
1993