NVIDIA is seeking an exceptional Manager, Deep Learning Inference Software, to lead a world-class engineering team advancing the state of AI model deployment. You will shape the software powering ...
NVIDIA is seeking an exceptional Manager, Deep Learning Inference Software, to lead a world-class engineering team advancing the state of AI model deployment. You will shape the software powering ...
Senior Deep Learning Performance Architect
Santa Clara, CA · On-site
$196K/yr
MS or PhD in Computer Science, Computer Engineering, Electrical Engineering or equivalent experience * 6+ years of meaningful work experience * Strong background in GPU or Deep Learning ASIC ...
Senior Deep Learning Performance Architect
Santa Clara, CA · On-site
$196K/yr
MS or PhD in Computer Science, Computer Engineering, Electrical Engineering or equivalent experience * 6+ years of meaningful work experience * Strong background in GPU or Deep Learning ASIC ...
Familiarity with deep learning accelerator architectures such as the GPU and hands-on experience with CUDA programming, kernel optimization, and workload profiling * Experience profiling and ...
Familiarity with deep learning accelerator architectures such as the GPU and hands-on experience with CUDA programming, kernel optimization, and workload profiling * Experience profiling and ...
Familiarity with deep learning accelerator architectures such as the GPU and hands-on experience with CUDA programming, kernel optimization, and workload profiling * Experience profiling and ...
Familiarity with deep learning accelerator architectures such as the GPU and hands-on experience with CUDA programming, kernel optimization, and workload profiling * Experience profiling and ...
MS or PhD in Computer Science, Computer Engineering, Electrical Engineering or equivalent experience * 6+ years of relevant meaningful work experience * Strong background in GPU or Deep Learning ASIC ...
MS or PhD in Computer Science, Computer Engineering, Electrical Engineering or equivalent experience * 6+ years of relevant meaningful work experience * Strong background in GPU or Deep Learning ASIC ...
MS or PhD in Computer Science, Computer Engineering, Electrical Engineering or equivalent experience * 6+ years of meaningful work experience * Strong background in GPU or Deep Learning ASIC ...
MS or PhD in Computer Science, Computer Engineering, Electrical Engineering or equivalent experience * 6+ years of meaningful work experience * Strong background in GPU or Deep Learning ASIC ...
(Senior) Software Engineer, Deep Learning
Fremont, CA · On-site
$140K - $280K/yr
... learning including data collection and analysis, evaluation and feature engineering. * Expertise in C++/Python. * Strong communication skills and team spirit. Preferred Experience * PhD in Deep ...
(Senior) Software Engineer, Deep Learning
Fremont, CA · On-site
$140K - $280K/yr
... learning including data collection and analysis, evaluation and feature engineering. * Expertise in C++/Python. * Strong communication skills and team spirit. Preferred Experience * PhD in Deep ...
We are now looking for a Senior GPU & Deep Learning Architect ... The NVIDIA GPU Architecture group is looking for world class architects and software developers to ...
We are now looking for a Senior GPU & Deep Learning Architect ... The NVIDIA GPU Architecture group is looking for world class architects and software developers to ...
Senior Deep Learning Performance Architect
Santa Clara, CA · On-site
$196K/yr
MS or PhD in Computer Science, Computer Engineering, Electrical Engineering or equivalent experience * 6+ years of relevant meaningful work experience * Strong background in GPU or Deep Learning ASIC ...
Senior Deep Learning Performance Architect
Santa Clara, CA · On-site
$196K/yr
MS or PhD in Computer Science, Computer Engineering, Electrical Engineering or equivalent experience * 6+ years of relevant meaningful work experience * Strong background in GPU or Deep Learning ASIC ...
(Senior) Software Engineer, Deep Learning
Fremont, CA · On-site
$140K - $280K/yr
... learning including data collection and analysis, evaluation and feature engineering. * Expertise in C++/Python. * Strong communication skills and team spirit. Preferred Experience * PhD in Deep ...
Quick apply
(Senior) Software Engineer, Deep Learning
Fremont, CA · On-site
$140K - $280K/yr
... learning including data collection and analysis, evaluation and feature engineering. * Expertise in C++/Python. * Strong communication skills and team spirit. Preferred Experience * PhD in Deep ...
(Senior) Software Engineer, Deep Learning
Fremont, CA · On-site
$140K - $280K/yr
... learning including data collection and analysis, evaluation and feature engineering. * Expertise in C++/Python. * Strong communication skills and team spirit. Preferred Experience * PhD in Deep ...
(Senior) Software Engineer, Deep Learning
Fremont, CA · On-site
$140K - $280K/yr
... learning including data collection and analysis, evaluation and feature engineering. * Expertise in C++/Python. * Strong communication skills and team spirit. Preferred Experience * PhD in Deep ...
Fluency in programming languages such as Python, C, C++. * Experience and familiarity with GPU ... Now, NVIDIA's GPU runs deep learning algorithms, simulating human intelligence, and acts as the ...
Fluency in programming languages such as Python, C, C++. * Experience and familiarity with GPU ... Now, NVIDIA's GPU runs deep learning algorithms, simulating human intelligence, and acts as the ...
Senior Deep Learning Frameworks CUDA Software Engineer
Santa Clara, CA · On-site
$143K - $189K/yr
We are looking for a motivated Deep Learning engineer to bring advanced CUDA features and Distributed Runtime technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will ...
Senior Deep Learning Frameworks CUDA Software Engineer
Santa Clara, CA · On-site
$143K - $189K/yr
We are looking for a motivated Deep Learning engineer to bring advanced CUDA features and Distributed Runtime technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will ...
Senior Deep Learning Performance Architect
Santa Clara, CA · On-site
$196K/yr
Fluency in programming languages such as Python, C, C++. * Experience and familiarity with GPU ... Now, NVIDIA's GPU runs deep learning algorithms, simulating human intelligence, and acts as the ...
Senior Deep Learning Performance Architect
Santa Clara, CA · On-site
$196K/yr
Fluency in programming languages such as Python, C, C++. * Experience and familiarity with GPU ... Now, NVIDIA's GPU runs deep learning algorithms, simulating human intelligence, and acts as the ...
We are now looking for a Senior GPU & Deep Learning Architect ... The NVIDIA GPU Architecture group is looking for world class architects and software developers to ...
We are now looking for a Senior GPU & Deep Learning Architect ... The NVIDIA GPU Architecture group is looking for world class architects and software developers to ...
Autonomy Engineer - Deep Learning Infrastructure
San Mateo, CA · On-site
$170K - $236K/yr
As a deep learning infrastructure engineer, you will be responsible for building and scaling the infrastructure that supports Skydio's DL and AI efforts. You will be working at the nexus of Skydio ...
Autonomy Engineer - Deep Learning Infrastructure
San Mateo, CA · On-site
$170K - $236K/yr
As a deep learning infrastructure engineer, you will be responsible for building and scaling the infrastructure that supports Skydio's DL and AI efforts. You will be working at the nexus of Skydio ...
We are looking for a motivated Deep Learning engineer to bring advanced CUDA features and Distributed Runtime technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will ...
We are looking for a motivated Deep Learning engineer to bring advanced CUDA features and Distributed Runtime technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will ...
We are looking for a motivated Deep Learning engineer to bring advanced communication technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will be working with the ...
We are looking for a motivated Deep Learning engineer to bring advanced communication technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will be working with the ...
Staff Software Engineer, Deep Learning Acceleration
San Francisco, CA · On-site
$189K - $274K/yr
As a Staff Software Engineer focusing on Deep Learning Acceleration at Aurora, you will play a pivotal role in enhancing the performance of Deep Learning networks utilized in our Autonomous Vehicle ...
Staff Software Engineer, Deep Learning Acceleration
San Francisco, CA · On-site
$189K - $274K/yr
As a Staff Software Engineer focusing on Deep Learning Acceleration at Aurora, you will play a pivotal role in enhancing the performance of Deep Learning networks utilized in our Autonomous Vehicle ...
We are looking for a motivated Deep Learning engineer to bring advanced communication technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will be working with the ...
We are looking for a motivated Deep Learning engineer to bring advanced communication technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will be working with the ...
Deep Learning Developer information
See Foster City, CA salary details
$21.01 - $24.50
2% of jobs
$24.50 - $27.99
0% of jobs
$27.99 - $31.48
0% of jobs
$31.48 - $34.97
12% of jobs
$34.97 - $38.46
11% of jobs
$38.65 is the 25th percentile. Wages below this are outliers.
$38.46 - $41.95
10% of jobs
The median wage is $45.03 / hr.
$41.95 - $45.44
18% of jobs
$45.44 - $48.93
16% of jobs
$50.35 is the 75th percentile. Wages above this are outliers.
$48.93 - $52.42
17% of jobs
$52.42 - $55.91
2% of jobs
$55.91 - $59.40
13% of jobs
$21
$44
$59
How much do deep learning developer jobs pay per hour?
What are the key skills and qualifications needed to thrive as a deep learning developer?
What is a deep learning developer?
What is the difference between Deep Learning Developer vs Machine Learning Engineer?
| Aspect | Deep Learning Developer | Machine Learning Engineer |
|---|---|---|
| Required Credentials | Bachelor's or Master's in CS, AI, or related; experience with neural networks | Bachelor's or Master's in CS, Data Science, or related; knowledge of algorithms |
| Work Environment | Research labs, AI startups, tech companies focusing on neural networks | Data-driven companies, software firms, industries applying machine learning |
| Industry Usage | Primarily in AI research, neural network development, deep learning projects | Broader application including predictive modeling, data analysis, and ML systems |
Deep Learning Developers specialize in neural networks and deep learning models, often working on AI research and complex algorithms. Machine Learning Engineers have a broader focus on developing, deploying, and maintaining machine learning models across various applications. While both roles require similar educational backgrounds, their focus areas and industry applications differ.
What are some common challenges deep learning developers face when deploying models to production environments?
Full-time
Posted 12 days ago
Nvidia rating
9.6
Based on 17 frontline employees who took The Breakroom Quiz
8th of 242 rated software companies
Job description
NVIDIA is seeking an exceptional Manager, Deep Learning Inference Software, to lead a world-class engineering team advancing the state of AI model deployment. You will shape the software powering today's most sophisticated AI systems - from large language models to multimodal generative AI - all accelerated on NVIDIA GPUs. The Deep Learning Inference team develops and optimizes open-source frameworks that make AI deployment scalable, efficient, and accessible - including SGLang, vLLM, and FlashInfer. Our work enables developers worldwide to harness NVIDIA accelerators for real-time inference at every scale, from datacenter clusters to edge devices.
What you'll be doing:
Lead, mentor, and scale a high-performing engineering team focused on deep learning inference and GPU-accelerated software.
Guide the strategy, roadmap, and execution of NVIDIA's OSS inference frameworks engineering.
Partner with internal compiler, libraries, and research teams to deliver end-to-end optimized inference pipelines across NVIDIA accelerators.
Oversee performance tuning, profiling, and optimization of large-scale models for LLM, multimodal, and generative AI applications.
Guide engineers in adopting best practices for CUDA, Triton, CUTLASS, and multi-GPU communications (NIXL, NCCL, NVSHMEM).
Represent the team in roadmap and planning discussions, ensuring alignment with NVIDIA's broader AI and software strategies.
Foster a culture of technical excellence, open collaboration, and continuous innovation.
What we need to see:
MS, PhD, or equivalent experience in Computer Science, Electrical/Computer Engineering, or a related field.
6+ overall years of software development experience, including 3+ years in technical leadership or engineering management.
Strong background in C/C++ software design and development; proficiency in Python is a plus.
Hands-on experience with GPU programming (CUDA, Triton, CUTLASS) and performance optimization.
Proven record of deploying or optimizing deep learning models in production environments.
Experience leading teams using Agile or collaborative software development practices.
Ways to Stand out from The Crowd:
Significant open-source contributions to deep learning or inference frameworks such as PyTorch, vLLM, SGLang, Triton, or TensorRT-LLM.
Deep understanding of multi-GPU communications (NIXL, NCCL, NVSHMEM) and distributed inference architectures.
Expertise in performance modeling, profiling, and system-level optimization across CPU and GPU platforms.
Proven ability to mentor engineers, guide architectural decisions, and deliver complex projects with measurable impact.
Publications, patents, or talks on LLM serving, model optimization, or GPU performance engineering.
With highly competitive salaries and a comprehensive benefits package, NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us and, our rapid growth means endless opportunities for career advancement.
If you're a passionate technical leader ready to shape the future of AI inference frameworks - and build the software that powers the world's most advanced models - we'd love to hear from you.
#LI-Hybrid
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 224,000 USD - 356,500 USD for Level 3, and 272,000 USD - 431,250 USD for Level 4.You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.About Nvidia
Sourced by ZipRecruiter
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.
Industry
Computer and electronic product manufacturing
Company size
10,000+ Employees
Headquarters location
Santa Clara, CA, US
Year founded
1993