1

Cuda Programming Jobs in California (NOW HIRING)

Strong CUDA programming skills with production kernel development * Deep understanding of GPU architecture (memory hierarchy, SMs, warps) * Track record of achieving significant performance ...

Strong CUDA programming skills with production kernel development * Deep understanding of GPU architecture (memory hierarchy, SMs, warps) * Track record of achieving significant performance ...

Senior Software Engineer - CUDA Driver

Santa Clara, CA · On-site

$143K - $189K/yr

Define forward-looking improvements to the CUDA APIs and programming model * Build and maintain performance and precision modeling * Write effective, maintainable, and well-tested code * Develop code ...

Preferred : • Experience with CUDA programming and real-time systems performance optimization. • Familiarity with scalable data engineering toolchains like Spark and Ray. • Strong publication ...

Senior Compiler Engineer - PVA

Santa Clara, CA · Hybrid

$122K - $168K/yr

NVIDIA has revolutionized parallel computing, and the development of the CUDA programming model has fueled its power. NVIDIA developed a powerful computing platform (PVA) focused on vision and deep ...

Showing results 21-40

Cuda Programming information

See California salary details

$27

$53

$80

How much do cuda programming jobs pay per hour?

As of Aug 19, 2026, the average hourly pay for cuda programming in California is $53.64, according to ZipRecruiter salary data. Most workers in this role earn between $43.41 and $62.64 per hour, depending on experience, location, and employer.

What is the difference between Cuda Programming vs GPU Developer?

AspectCuda ProgrammingGPU Developer
Required CredentialsKnowledge of CUDA, C/C++, parallel computingKnowledge of GPU architecture, CUDA, OpenCL, C/C++
Work EnvironmentHigh-performance computing, scientific research, AIGraphics, gaming, scientific visualization, AI
Industry UsageTech companies, research labs, AI firmsGaming, entertainment, tech, research

While Cuda Programming focuses specifically on writing code using NVIDIA's CUDA platform for parallel processing, GPU Developers have a broader role that includes designing, optimizing, and implementing GPU-based solutions across various platforms and technologies. Both roles require knowledge of GPU architecture and programming languages like C/C++, but GPU Developers often work on a wider range of applications beyond CUDA-specific projects.

Are CUDA programmers in demand?

CUDA programmers are in high demand due to the growing need for high-performance computing in fields like artificial intelligence, scientific research, and data processing. Skills in parallel programming, GPU architecture, and CUDA toolkit are highly valued, and job opportunities are expected to grow as industries adopt GPU acceleration for complex tasks.

What does a CUDA programming developer do?

A CUDA programming developer writes software that leverages NVIDIA's CUDA platform to perform parallel processing on GPUs, optimizing computational tasks such as scientific simulations, machine learning, and image processing. They typically work with C++ and CUDA-specific libraries, debugging and optimizing code for high performance in environments that require intensive data processing.

What job categories do people searching Cuda Programming jobs in California look for?

The top searched job categories for Cuda Programming jobs in California are:

What cities in California are hiring for Cuda Programming jobs?

Cities in California with the most Cuda Programming job openings:

Infographic showing various Cuda Programming job openings in California as of August 2026, with employment types broken down into 92% Full Time, and 8% Contract. Highlights an 76% In-person, and 24% Remote job distribution, with an average salary of $111,580 per year, or $53.6 per hour.

GPU Performance Engineer

Genmo

San Francisco, CA • On-site

Full-time

Re-posted 16 days ago


Job description

We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.
We're seeking a GPU Performance Engineer to squeeze every last FLOP from our H100 infrastructure and optimize our model serving stack to its absolute limits.
The Role
You'll be our performance optimization expert, using advanced profiling tools to identify bottlenecks and implementing solutions that achieve 5-10x speedups. From writing custom CUDA kernels to eliminating cold start latency, you'll ensure our infrastructure delivers world-class performance. This role is perfect for someone who gets excited about microsecond optimizations and pushing hardware to its theoretical limits.
Key Responsibilities
  • Profile and optimize GPU workloads using Nsight Systems, nvprof, and custom instrumentation
  • Write high-performance CUDA and Triton kernels for critical model operations
  • Optimize cold start latency from seconds to milliseconds for our serving infrastructure
  • Tune memory access patterns, kernel fusion, and GPU utilization
  • Collaborate with ML engineers to optimize model implementations
  • Debug performance issues across the full stack from application to hardware
  • Implement custom memory pooling and allocation strategies
  • Share optimization techniques and build performance culture across teams

Qualifications
  • Bachelor's or Master's degree in Computer Science, Electrical Engineering, or related field
  • 5+ years systems programming experience with 3+ years focused on GPU optimization
  • Expert proficiency with GPU profiling tools (Nsight Systems, nvprof)
  • Strong CUDA programming skills with production kernel development
  • Deep understanding of GPU architecture (memory hierarchy, SMs, warps)
  • Track record of achieving significant performance improvements (5-10x)
  • Experience with Python and C++ in production environments

We Value
  • Experience with Triton kernel development
  • Knowledge of CUTLASS or similar high-performance libraries
  • Background in ML-specific optimizations (attention, transformers)
  • RDMA/InfiniBand optimization experience
  • Contributions to GPU libraries or frameworks
  • Low-level debugging skills (PTX/SASS reading)

Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.