1

Cuda Programmer Jobs in Texas (NOW HIRING)

GPU Programmer - Remote Job Type: Contractor Location: Remote Job Overview We are seeking ... Design, implement, and optimize GPU software using CUDA, WebGPU, or GLSL . * Profile and optimize ...

GPU Programmer - Remote

Dallas, TX · Remote

$60 - $85/hr

GPU Programmer - Remote Job Type: Contractor Location: Remote Job Overview We are seeking ... Design, implement, and optimize GPU software using CUDA, WebGPU, or GLSL . * Profile and optimize ...

GPU Programmer - Remote

Austin, TX · Remote

$60 - $85/hr

GPU Programmer - Remote Job Type: Contractor Location: Remote Job Overview We are seeking ... Design, implement, and optimize GPU software using CUDA, WebGPU, or GLSL . * Profile and optimize ...

Systems Software Engineer, (Linux, CUDA)

Richardson, TX · On-site

$157K - $186K/yr

Experience with GPU programming and runtimes (e.g., CUDA, SYCL, OpenCL, PyTorch) and familiarity with GPU memory hierarchies. * Strong development skills in C, C++, Python, and shell scripting, with ...

Senior Compiler Engineer

Austin, TX · On-site

$103K - $142K/yr

We are redefining how developers write high-performance GPU software by bringing the safety, expressiveness, and modern tooling of Rust to native GPU and CUDA development. On this team, you will ...

Senior Compiler Engineer

Austin, TX · On-site

$103K - $142K/yr

We are redefining how developers write high-performance GPU software by bringing the safety, expressiveness, and modern tooling of Rust to native GPU and CUDA development. On this team, you will ...

Senior Compiler Engineer - Rust GPU

Austin, TX · On-site

$121K - $160K/yr

We are redefining how developers write high-performance GPU software by bringing the safety, expressiveness, and modern tooling of Rust to native GPU and CUDA development. On this team, you will ...

Senior Compiler Engineer - Rust GPU

Austin, TX · On-site

$121K - $160K/yr

We are redefining how developers write high-performance GPU software by bringing the safety, expressiveness, and modern tooling of Rust to native GPU and CUDA development. On this team, you will ...

Senior GPU Compiler Development Engineer

Austin, TX · On-site

$121K - $160K/yr

Work with NVIDIA GPU Architecture and CUDA Programming model teams to build abstractions to expose new GPU features in portable and performant ways in PTX ISA. PTX Compiler (PTXAS) apart from ...

Showing results 21-40

Cuda Programmer information

See Texas salary details

$11

$36

$64

How much do cuda programmer jobs pay per hour?

As of Sep 7, 2026, the average hourly pay for cuda programmer in Texas is $36.83, according to ZipRecruiter salary data. Most workers in this role earn between $23.94 and $47.93 per hour, depending on experience, location, and employer.

What is a cuda programmer?

A CUDA Programmer develops high-performance parallel computing applications using NVIDIA's CUDA (Compute Unified Device Architecture) framework. They optimize algorithms to run efficiently on GPUs, accelerating tasks such as machine learning, scientific simulations, and real-time data processing. This role requires proficiency in C/C++, an understanding of GPU architectures, and experience with parallel computing concepts to maximize performance.

What are the key skills and qualifications needed to thrive as a cuda programmer?

To thrive as a Cuda Programmer, you need strong programming skills in C/C++ and parallel computing, with a solid understanding of GPU architectures and CUDA development. Familiarity with CUDA libraries, performance profiling tools, and platforms like NVIDIA Nsight or Visual Studio is often required, while certifications from NVIDIA can be advantageous. Problem-solving abilities, attention to detail, and effective teamwork and communication skills help set candidates apart. These competencies ensure you can optimize complex algorithms, work efficiently on high-performance computing projects, and collaborate smoothly with multidisciplinary teams.

What are the most common challenges faced by cuda programmers in their daily work?

Cuda Programmers often encounter challenges related to optimizing code performance and efficiently managing memory on GPU architectures. Debugging and profiling can be complex, as issues may arise from both the code and hardware-specific elements, requiring close attention to parallelization and bottlenecks. Collaboration is key, as you’ll typically work closely with software engineers, data scientists, or researchers to integrate and optimize code for specialized workflows. Successfully navigating these challenges helps drive significant performance improvements and innovation in high-performance computing applications.

What are the most commonly searched types of Cuda Programmer jobs in Texas?

The most popular types of Cuda Programmer jobs in Texas are:

What are popular job titles related to Cuda Programmer jobs in Texas?

For Cuda Programmer jobs in Texas, the most frequently searched job titles are:

What job categories do people searching Cuda Programmer jobs in Texas look for?

The top searched job categories for Cuda Programmer jobs in Texas are:

Infographic showing various Cuda Programmer job openings in Texas as of August 2026, with employment types broken down into 81% Full Time, 7% Part Time, 10% Contract, and 2% Nights. Highlights an 89% Physical, 3% Hybrid, and 8% Remote job distribution, with an average salary of $76,614 per year, or $36.8 per hour.

Senior Deep Learning Frameworks CUDA Software Engineer

Nvidia

Austin, TX • On-site

$121K - $160K/yr

Full-time

Re-posted 24 days ago


Nvidia rating

9.6

Company rating: 9.6 out of 10

Based on 18 frontline employees who took The Breakroom Quiz

6th of 247 rated software companies


Job description

NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction inventions from artificial intelligence to autonomous cars.

We are looking for a motivated Deep Learning engineer to bring advanced CUDA features and Distributed Runtime technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will be working with the team that created core CUDA features and runtimes for scaling Deep Learning and HPC applications. Your customers will have diverse multi-GPU demands, ranging from training on scales up to 100K GPUs to inference down at microsecond latency.

CUDA features improve both productivity and performance of AI applications. Your work in AI toolkits will accelerate enabling those for the community. This is an outstanding opportunity for someone with an AI background to advance the state of the art in this space.

Are you ready to contribute to the development of innovative technologies and help realize NVIDIA's vision. What you will be doing: Integrate new CUDA features and Runtime abstractions in AI frameworks: from PoC to performance analysis to production Perform deep analysis of AI workloads and frameworks to identify requirements and opportunities to innovate in the lower layers of the stack. Collaborate hands-on with teams working on the latest AI models.

Own and drive improvements in the AI Compiler-Runtime interface to build speed-of-light multi-GPU multi-node solutions. Design fault-tolerant and elastic solutions for large-scale or dynamic AI workloads. Influence the roadmap of core CUDA to facilitate building next-gen DL frameworks.

Collaborate with a very dynamic team across multiple time zones. Collaborate closely with AI researchers, HW and SW architects, kernel and compiler authors and CUDA driver experts to co-design systems and frameworks that enhance performance and programmability. Develop exploratory tools and runtime systems to profile and accelerate new paradigms in deep learning.

Write clean, effective, and maintainable code, ensuring exploratory prototypes can smoothly transition into open-source releases, upstream framework integrations, internal tools, or closed-source commercial products. What we need to see: BS, MS, or PhD degree in Computer Science, Computer Engineering, Electrical Engineering, or related field (or equivalent experience). 8+ years of relevant industry experience or equivalent academic experience after completed degree.

Development experience with Deep Learning Frameworks such PyTorch, JAX, and Inference Engines such as TRT-LLM, vLLM, SGLang Rapid prototyping and development with Python, C++, CUDA or related DSLs Solid grasp of AI models, parallelisms, and/or compiler technologies (e.g. torch.compile) Experience conducting performance benchmarking on AI clusters. Familiarity with at least one performance profiler toolchain (PyTorch profiler, NVIDIA Nsight Systems) Understanding of HPC/AI communication concepts Good understanding of computer system architecture, HW-SW interactions and operating systems principles (aka systems software fundamentals) Adaptability and passion to learn new frameworks and tools Flexibility to work and communicate effectively across different teams and timezones Ways to stand out from the crowd: Deep expertise in the performance internals and execution graphs of major deep learning autograd, training and inference frameworks (e.g., PyTorch, JAX, TensorRT, vLLM, sgLang, Nemo, Megatron, MaxText, etc.)

Hands-on experience with CUDA, specific communication libraries (e.g., NCCL, MPI, UCX) and distributed machine learning techniques (e.g., pipeline parallelism, tensor parallelism). Expertise in one or more of these areas: Training, Distributed inference, MoE, Reinforcement Learning, kernel authoring (on CUDA, Triton, cuTe, etc). Background in deep learning compilers, both graph-level and codegen (e.g., Triton, XLA, torch compile) Experience with programming for compute & communication overlap in distributed runtime Your base salary will be determined based on your location, experience, and the pay of employees in similar positions

The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. You will also be eligible for equity and benefits. Applications for this job will be accepted at least until September 4, 2026.

This posting is for an existing vacancy. NVIDIA uses AI tools in its recruiting processes. NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer.

As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.


What Nvidia employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Nvidia logo

About Nvidia

Sourced by ZipRecruiter

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.

Industry

Computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

Santa Clara, CA, US