1

Cuda Library Software Engineer Jobs (NOW HIRING)

$122K - $161K/yr

With flexible I/O, blazing-fast merging, aggregating, and filtering, our library serves diverse ... We use the latest tools in modern C++ and CUDA to produce software with elegant design, broad ...

$90K - $119K/yr

With flexible I/O, blazing-fast merging, aggregating, and filtering, our library serves diverse ... We use the latest tools in modern C++ and CUDA to produce software with elegant design, broad ...

Senior Software Engineer - CUDA Driver

Santa Clara, CA · On-site

$143K - $189K/yr

You will join a versatile software engineering team that delivers innovative software features to ... Strong background in parallel computing, preferably writing CUDA programs or CUDA-based libraries.

Showing results 21-40

Cuda Library Software Engineer information

See salary details

$96

$112

$128

How much do cuda library software engineer jobs pay per hour?

As of Sep 11, 2026, the average hourly pay for cuda library software engineer in the United States is $112.98, according to ZipRecruiter salary data. Most workers in this role earn between $104.57 and $121.39 per hour, depending on experience, location, and employer.

What are popular job titles related to Cuda Library Software Engineer jobs?

For Cuda Library Software Engineer jobs, the most frequently searched job titles are:

Senior Software Engineer, CUDA Python Core Libraries

Remote

Nvidia
Computer and Electronic Product Manufacturing • 10K+ employees

$124K - $167K/yr

Full-time

Posted 26 days ago


Nvidia rating

9.6

Company rating: 9.6 out of 10

Based on 18 frontline employees who took The Breakroom Quiz


Job description

NVIDIA's accelerated computing platform is foundational to modern HPC and AI. At the center of this platform are CUDA Core Libraries that enable developers to build fast, reliable, and scalable GPU-accelerated software. We are hiring a Senior Software Engineer to advance the Python experience for CUDA Core Libraries.

You will build Pythonic APIs, language bindings, algorithms, and runtime infrastructure on top of native C/C++ foundations. You will join the team building the foundational libraries, algorithms, and language/runtime infrastructure that make CUDA a speed-of-light experience for developers and AI coding agents alike. What you'll be doing: Design and implement idiomatic Python APIs and bindings for foundational CUDA capabilities and GPU algorithms.

Develop and integrate the native C/C++ components that support Python-facing functionality. Define reliable and efficient interoperability boundaries between Python, C/C++, Rust, and other languages. Develop high-performance interfaces that minimize Python and native-language integration overhead.

Own features throughout their lifecycle: design, implementation, testing, profiling, benchmarking, documentation, release, and long-term maintenance. Improve the Python developer experience through typing, packaging, examples, diagnostics, continuous integration, and compatibility testing. Collaborate with C/C++, Rust, compiler, and runtime engineers on shared architecture and API decisions.

Work directly with users to investigate correctness, usability, compatibility, and performance issues. What we need to see: BS, MS, or PhD in Computer Science, Computer Engineering, or a related field, or equivalent experience. 8+ years of relevant software-development experience.

Strong production programming skills in both Python and C/C++; both are required for this role. Experience building Python interfaces to native or systems-level software. Understanding of systems software concepts, performance, concurrency, and API design.

Practical experience with parallel, heterogeneous, or GPU programming. Experience developing production software or widely used libraries, including testing, profiling, benchmarking, packaging, and code review. Ability to work independently, define project scope, and drive complex work to completion.

Clear written communication skills for API specifications, technical designs, and user documentation. Comfort working in large codebases spanning Python, C/C++, build systems, packaging, and continuous-integration infrastructure. Ways to stand out from the crowd: Strong understanding of CPU/GPU architecture and performance optimization, with hands-on experience in GPU-accelerated stacks (CUDA C++/Python, PyTorch, JAX, Numba, CuPy, or similar).

Proficiency with modern C++ and GPU libraries such as Thrust, CUB, and libcudacxx. Experience with compiler infrastructure and tooling, including LLVM, Clang, or MLIR. Expertise in designing low-overhead interoperability between Python and native languages, including exposure to Rust in mixed-language stacks.

Demonstrated interest in developer tools, library design, and improving developer productivity. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits. Applications for this job will be accepted at least until September 4, 2026. This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.


What Nvidia employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Nvidia logo

About Nvidia

Sourced by ZipRecruiter

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.

Industry

Computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

Santa Clara, CA, US