Familiarity with CUDA programming and/or GPUs. * Experience with HPC or large-scale computing environments. Your base salary will be determined based on your location, experience, and the pay of ...
Familiarity with CUDA programming and/or GPUs. * Experience with HPC or large-scale computing environments. Your base salary will be determined based on your location, experience, and the pay of ...
Senior AI Software Architect - Autonomous Systems
Austin, TX · On-site
$147K/yr
Deep expertise in ROCm or CUDA programming, GPU kernel optimization, and GPU memory management for accelerating AI inference and training workloads. * System Performance Analysis: Expertise in ...
Senior AI Software Architect - Autonomous Systems
Austin, TX · On-site
$147K/yr
Deep expertise in ROCm or CUDA programming, GPU kernel optimization, and GPU memory management for accelerating AI inference and training workloads. * System Performance Analysis: Expertise in ...
Senior AI Software Architect - Autonomous Systems
Austin, TX · Hybrid
$128K - $174K/yr
Deep expertise in ROCm or CUDA programming, GPU kernel optimization, and GPU memory management for accelerating AI inference and training workloads. * System Performance Analysis: Expertise in ...
Senior AI Software Architect - Autonomous Systems
Austin, TX · Hybrid
$128K - $174K/yr
Deep expertise in ROCm or CUDA programming, GPU kernel optimization, and GPU memory management for accelerating AI inference and training workloads. * System Performance Analysis: Expertise in ...
CUDA programming and NVIDIA GPU architecture expertise. * Proved experience influencing product strategy and technical roadmap at a senior level. * Major open-source contributions. With competitive ...
CUDA programming and NVIDIA GPU architecture expertise. * Proved experience influencing product strategy and technical roadmap at a senior level. * Major open-source contributions. With competitive ...
Proficient with NVIDIA GPUs, CUDA programming, and NCCL, including performance benchmarking via MLPerf. * Hardware & Storage Engineering: Deep familiarity with storage hardware (HDDs, SSDs, NVMe ...
Proficient with NVIDIA GPUs, CUDA programming, and NCCL, including performance benchmarking via MLPerf. * Hardware & Storage Engineering: Deep familiarity with storage hardware (HDDs, SSDs, NVMe ...
Senior HPC Cluster Engineer
$103K - $142K/yr
Background with NVIDIA GPUs, CUDA Programming, NCCL and MLPerf benchmarking. * Experience supporting EDA workloads and tools. * Familiarity with High-Speed Networking pertaining to HPC including ...
Senior HPC Cluster Engineer
$103K - $142K/yr
Background with NVIDIA GPUs, CUDA Programming, NCCL and MLPerf benchmarking. * Experience supporting EDA workloads and tools. * Familiarity with High-Speed Networking pertaining to HPC including ...
A grasp of the CUDA programming model and experience employing GPU profiling tools like NVIDIA Nsight Systems/Compute to address PCIe bottlenecks and kernel stalls. * Extensive knowledge of profiling ...
A grasp of the CUDA programming model and experience employing GPU profiling tools like NVIDIA Nsight Systems/Compute to address PCIe bottlenecks and kernel stalls. * Extensive knowledge of profiling ...
Principal Architect, AI Networking
Austin, TX · On-site
$272 - $431.25/hr
CUDA programming and NVIDIA GPU architecture expertise.* Proved experience influencing product strategy and technical roadmap at a senior level.* Major open-source contributions.With competitive ...
Principal Architect, AI Networking
Austin, TX · On-site
$272 - $431.25/hr
CUDA programming and NVIDIA GPU architecture expertise.* Proved experience influencing product strategy and technical roadmap at a senior level.* Major open-source contributions.With competitive ...
Senior Software Architect - Deep Learning and HPC Communications (Austin)
Austin, TX · On-site
$128K - $174K/yr
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high‑performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such as PyTorch, TensorFlow ...
Senior Software Architect - Deep Learning and HPC Communications (Austin)
Austin, TX · On-site
$128K - $174K/yr
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high‑performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such as PyTorch, TensorFlow ...
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high-performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such PyTorch, TensorFlow, etc.
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high-performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such PyTorch, TensorFlow, etc.
Senior Software Architect - Deep Learning and HPC Communications
Austin, TX · On-site
$128K - $174K/yr
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high-performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such PyTorch, TensorFlow, etc.
Senior Software Architect - Deep Learning and HPC Communications
Austin, TX · On-site
$128K - $174K/yr
Experience with CUDA programming and NVIDIA GPUs. Knowledge of high-performance networks like InfiniBand, RoCE, NVLink, etc. * Experience with Deep Learning Frameworks such PyTorch, TensorFlow, etc.
Expert knowledge of the NVIDIA GPU memory hierarchy (HBM3e/HBM4, L2 cache) and CUDA programming models. Ways to Stand Out from the Crowd: * Framework Development:Hands-on experience developing within ...
Expert knowledge of the NVIDIA GPU memory hierarchy (HBM3e/HBM4, L2 cache) and CUDA programming models. Ways to Stand Out from the Crowd: * Framework Development:Hands-on experience developing within ...
Staff Engineer, Inference Optimizations
Austin, TX · Remote
$191K - $239K/yr
Excellent system design skills, particularly related to low-level GPU programming - optimization ... Expert-level Triton or CUDA. If you've contributed to the Triton compiler or wrote custom CUDA ...
Staff Engineer, Inference Optimizations
Austin, TX · Remote
$191K - $239K/yr
Excellent system design skills, particularly related to low-level GPU programming - optimization ... Expert-level Triton or CUDA. If you've contributed to the Triton compiler or wrote custom CUDA ...
Senior Software Engineer, CUTLASS Kernels
$121K - $160K/yr
Experience with CUDA, OpenCL, HIP, SYCL, Mojo, Pallas, Triton, Mosaic, Halide, or any general-purpose or domain-specific programming language targeting highly parallel accelerators. * Deep ...
Senior Software Engineer, CUTLASS Kernels
$121K - $160K/yr
Experience with CUDA, OpenCL, HIP, SYCL, Mojo, Pallas, Triton, Mosaic, Halide, or any general-purpose or domain-specific programming language targeting highly parallel accelerators. * Deep ...
Principal Software Engineer, Profiling Services
$133K - $179K/yr
You will lead the architecture and hands-on delivery across system software, drivers, and CUDA to ... Set technical direction for an engineering team; mentor engineers, drive technical planning to ...
Principal Software Engineer, Profiling Services
$133K - $179K/yr
You will lead the architecture and hands-on delivery across system software, drivers, and CUDA to ... Set technical direction for an engineering team; mentor engineers, drive technical planning to ...
Principal Software Engineer, Profiling Services
Austin, TX · On-site
$133K - $179K/yr
They are seeking a Principal Software Engineer to design and implement a low-overhead GPU profiling ... of CUDA and GPU architecture, including runtime/driver APIs, CUDA streams/graphs, and kernel ...
Principal Software Engineer, Profiling Services
Austin, TX · On-site
$133K - $179K/yr
They are seeking a Principal Software Engineer to design and implement a low-overhead GPU profiling ... of CUDA and GPU architecture, including runtime/driver APIs, CUDA streams/graphs, and kernel ...
Senior Deep Learning Engineer
Austin, TX · On-site +1
$130K - $180K/yr
Experience in embedded or low-level programming * Knowledge of CUDA/OpenGL * Experience deploying neural networks in production * Familiarity with model compression techniques like quantization ...
Quick apply
Senior Deep Learning Engineer
Austin, TX · On-site +1
$130K - $180K/yr
Experience in embedded or low-level programming * Knowledge of CUDA/OpenGL * Experience deploying neural networks in production * Familiarity with model compression techniques like quantization ...
Principal Software Engineer - Control Systems
Austin, TX · On-site
$165 - $200/hr
Rapidly prototype and iterate using Rust, Python, C/C++, and CUDA * Collaborate with quantum theorists, hardware engineers, and platform software teams to translate quantum control flows into fast ...
Principal Software Engineer - Control Systems
Austin, TX · On-site
$165 - $200/hr
Rapidly prototype and iterate using Rust, Python, C/C++, and CUDA * Collaborate with quantum theorists, hardware engineers, and platform software teams to translate quantum control flows into fast ...
Principal Software Engineer - Control Systems (Austin)
Austin, TX · On-site
$165K - $200K/yr
Rapidly prototype and iterate using Rust, Python, C/C++, and CUDA * Collaborate with quantum theorists, hardware engineers, and platform software teams to translate quantum control flows into fast ...
Principal Software Engineer - Control Systems (Austin)
Austin, TX · On-site
$165K - $200K/yr
Rapidly prototype and iterate using Rust, Python, C/C++, and CUDA * Collaborate with quantum theorists, hardware engineers, and platform software teams to translate quantum control flows into fast ...
Senior Software Engineer - TensorRT Edge-LLM
Austin, TX · Hybrid
$121K - $160K/yr
Contribute to CUDA kernel and operator development for critical transformer components such as ... Proficient programming ability with modern C++ (C++11/14/17 and beyond). * Familiarity with popular ...
Senior Software Engineer - TensorRT Edge-LLM
Austin, TX · Hybrid
$121K - $160K/yr
Contribute to CUDA kernel and operator development for critical transformer components such as ... Proficient programming ability with modern C++ (C++11/14/17 and beyond). * Familiarity with popular ...
Cuda Programming information
See Austin, TX salary details
$27.67 - $32.54
5% of jobs
$32.54 - $37.42
10% of jobs
$37.42 - $42.30
9% of jobs
$43.35 is the 25th percentile. Wages below this are outliers.
$42.30 - $47.18
7% of jobs
$47.18 - $52.06
15% of jobs
The median wage is $53.56 / hr.
$52.06 - $56.94
14% of jobs
$61.36 is the 75th percentile. Wages above this are outliers.
$56.94 - $61.81
17% of jobs
$61.81 - $66.69
14% of jobs
$66.69 - $71.57
6% of jobs
$71.57 - $76.45
3% of jobs
$76.45 - $81.33
0% of jobs
$27
$53
$81
How much do cuda programming jobs pay per hour?
Are CUDA programmers in demand?
What is the difference between Cuda Programming vs GPU Developer?
| Aspect | Cuda Programming | GPU Developer |
|---|---|---|
| Required Credentials | Knowledge of CUDA, C/C++, parallel computing | Knowledge of GPU architecture, CUDA, OpenCL, C/C++ |
| Work Environment | High-performance computing, scientific research, AI | Graphics, gaming, scientific visualization, AI |
| Industry Usage | Tech companies, research labs, AI firms | Gaming, entertainment, tech, research |
While Cuda Programming focuses specifically on writing code using NVIDIA's CUDA platform for parallel processing, GPU Developers have a broader role that includes designing, optimizing, and implementing GPU-based solutions across various platforms and technologies. Both roles require knowledge of GPU architecture and programming languages like C/C++, but GPU Developers often work on a wider range of applications beyond CUDA-specific projects.
What does a CUDA programmer do?

Nvidia rating
9.6
Based on 17 frontline employees who took The Breakroom Quiz
8th of 242 rated software companies
Job description
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology-and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our tightly coupled CPU, GPU and DPU technology acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent. As an NVIDIAN, you'll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world!
NVIDIA is searching for a highly motivated, technical engineer to join the Tegra system-on-chip (SoC) software organization. You will work on key aspects of our ARM SW ecosystem and system software architecture. With a targeted charter to enable best-in-class datacenter-scale performance and efficiency for our next generation of datacenter products, including CPUs and CPU+GPU Superchips.
What you will be doing:
Design, develop, test, and optimize software for our next-generation SoCs. In both pre-silicon and post-silicon phases of execution.
Review architectural performance bottlenecks for various system wide work loads. Identify HW/SW policies to drive performance and performance/watt leadership.
Using strong communication skills, build and drive architecture, analysis documents and communications to internal and/or external audiences about our technology.
Competitive analysis comparing uArchitecture & workload performance metrics on NVIDIA's ARM SoCs against emerging processors from other silicon vendors.
Influence and drive full-stack adoption of performance optimizations and best practices across NVIDIA SW products & OSS SDKs
What we need to see:
BS or MS degree in Computer Engineering, Computer Science, or related degree (or equivalent experience).
6+ years of relevant computer architecture or SW development experience.
Proven leadership skills and strong ownership on past projects.
Hands on technical experience and demonstrated excellence in an environment with complex software and hardware designs.
Strong understanding of multicore hardware, operating systems design, concurrency, virtual memory, caching, interrupts, device drivers and real-time programming.
Strong stills in performance analysis, data analysis and performance optimization.
Strong use of linux perf tool, application/library performance optimizations.
Ways to stand out from the crowd:
Deep expertise in ARM architecture and SW ecosystem.
Proficient in analyzing, debugging and tuning performance of complex system software stacks.
Experience with CPU server system workloads and performance analysis.
Familiarity with CUDA programming and/or GPUs.
Experience with HPC or large-scale computing environments.
You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.About Nvidia
Sourced by ZipRecruiter
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.
Industry
Computer and electronic product manufacturing
Company size
10,000+ Employees
Headquarters location
Santa Clara, CA, US
Year founded
1993