This internship provides hands-on experience in low-level GPU performance analysis, kernel ... Experience with CUDA programming and GPU kernel development. * Understanding of NVIDIA GPU ...
This internship provides hands-on experience in low-level GPU performance analysis, kernel ... Experience with CUDA programming and GPU kernel development. * Understanding of NVIDIA GPU ...
Scientific Software Developer Paid Co-op/Internship
Albuquerque, NM · On-site
$18.75 - $24.50/hr
Selected Interns or Co-op students will support object-oriented C++ software development in the ... Supercomputing, CUDA, OpenMP, MPI, threads, GPU * Image processing, imagery analysis, computer ...
Scientific Software Developer Paid Co-op/Internship
Albuquerque, NM · On-site
$18.75 - $24.50/hr
Selected Interns or Co-op students will support object-oriented C++ software development in the ... Supercomputing, CUDA, OpenMP, MPI, threads, GPU * Image processing, imagery analysis, computer ...
Scientific Software Developer Paid Co-op/Internship
Albuquerque, NM · On-site
$17.50 - $23/hr
Selected Interns or Co-op students will support object-oriented C++ software development in the ... Supercomputing, CUDA, OpenMP, MPI, threads, GPU * Image processing, imagery analysis, computer ...
Quick apply
Scientific Software Developer Paid Co-op/Internship
Albuquerque, NM · On-site
$17.50 - $23/hr
Selected Interns or Co-op students will support object-oriented C++ software development in the ... Supercomputing, CUDA, OpenMP, MPI, threads, GPU * Image processing, imagery analysis, computer ...
... Internship program. In addition to technical acumen, we highly value leadership skills as we ... Experience with GPGPU development using CUDA, OpenCL, or OpenGL Compute * Experience with machine ...
... Internship program. In addition to technical acumen, we highly value leadership skills as we ... Experience with GPGPU development using CUDA, OpenCL, or OpenGL Compute * Experience with machine ...
Senior Algorithm Engineer (Image Processing & CUDA C++)
Milpitas, CA · On-site
$159K - $271K/yr
Experience in CUDA C++ and GPU-accelerated computing is required * Substantial work experience in ... Interns are eligible for some of the benefits listed. Our pay ranges are determined by role, level ...
Senior Algorithm Engineer (Image Processing & CUDA C++)
Milpitas, CA · On-site
$159K - $271K/yr
Experience in CUDA C++ and GPU-accelerated computing is required * Substantial work experience in ... Interns are eligible for some of the benefits listed. Our pay ranges are determined by role, level ...
Senior Algorithm Engineer (Image Processing & CUDA C++)
Milpitas, CA · On-site
$159K - $271K/yr
Experience in CUDA C++ and GPU-accelerated computing is required * Substantial work experience in ... Interns are eligible for some of the benefits listed. Our pay ranges are determined by role, level ...
Senior Algorithm Engineer (Image Processing & CUDA C++)
Milpitas, CA · On-site
$159K - $271K/yr
Experience in CUDA C++ and GPU-accelerated computing is required * Substantial work experience in ... Interns are eligible for some of the benefits listed. Our pay ranges are determined by role, level ...
Senior Algorithm Engineer (Image Processing & CUDA C++)
Milpitas, CA · On-site
$159K - $271K/yr
Experience in CUDA C++ and GPU-accelerated computing is required * Substantial work experience in ... Interns are eligible for some of the benefits listed. Our pay ranges are determined by role, level ...
Senior Algorithm Engineer (Image Processing & CUDA C++)
Milpitas, CA · On-site
$159K - $271K/yr
Experience in CUDA C++ and GPU-accelerated computing is required * Substantial work experience in ... Interns are eligible for some of the benefits listed. Our pay ranges are determined by role, level ...
Senior Algorithm Engineer (Image Processing & CUDA C++)
Milpitas, CA · On-site
$159K - $271K/yr
Experience in CUDA C++ and GPU-accelerated computing is required * Substantial work experience in ... Interns are eligible for some of the benefits listed. Our pay ranges are determined by role, level ...
Senior Algorithm Engineer (Image Processing & CUDA C++)
Milpitas, CA · On-site
$159K - $271K/yr
Experience in CUDA C++ and GPU-accelerated computing is required * Substantial work experience in ... Interns are eligible for some of the benefits listed. Our pay ranges are determined by role, level ...
Legally authorized to do an internship in the United States for either 3 months or 6 months. Nice to have * Experience with sensor drivers, CUDA, 3D geometry, or multi-sensor fusion.
Legally authorized to do an internship in the United States for either 3 months or 6 months. Nice to have * Experience with sensor drivers, CUDA, 3D geometry, or multi-sensor fusion.
Senior Research Scientist, Efficient Deep Learning
Santa Clara, CA · On-site
$115K - $147K/yr
Mentor interns. • Work with product groups to transfer technology. Qualifications : Required ... C++ and parallel programming (e.g., CUDA). • Hands-on experience with large-scale model training ...
Senior Research Scientist, Efficient Deep Learning
Santa Clara, CA · On-site
$115K - $147K/yr
Mentor interns. • Work with product groups to transfer technology. Qualifications : Required ... C++ and parallel programming (e.g., CUDA). • Hands-on experience with large-scale model training ...
Algorithm engineer interns at Ambarella are responsible for developing highly efficient algorithms ... Strong programming skills in Python, C/C++, CUDA. * Excellent communication skills.
Algorithm engineer interns at Ambarella are responsible for developing highly efficient algorithms ... Strong programming skills in Python, C/C++, CUDA. * Excellent communication skills.
... Internship program. In addition to technical acumen, we highly value leadership skills as we ... Experience with GPGPU development using CUDA, OpenCL, or OpenGL Compute * Experience with machine ...
... Internship program. In addition to technical acumen, we highly value leadership skills as we ... Experience with GPGPU development using CUDA, OpenCL, or OpenGL Compute * Experience with machine ...
... Internship program. In addition to technical acumen, we highly value leadership skills as we ... Experience with GPGPU development using CUDA, OpenCL, or OpenGL Compute * Experience with machine ...
... Internship program. In addition to technical acumen, we highly value leadership skills as we ... Experience with GPGPU development using CUDA, OpenCL, or OpenGL Compute * Experience with machine ...
Internship or research experience with LLM inference, ML systems, or model serving * Contributions to open-source inference frameworks (vLLM, SGLang, TensorRT-LLM, etc.) * CUDA / Triton kernel work ...
Internship or research experience with LLM inference, ML systems, or model serving * Contributions to open-source inference frameworks (vLLM, SGLang, TensorRT-LLM, etc.) * CUDA / Triton kernel work ...
Member of Technical Staff - Model Optimization and Inference (New Grad)
Seattle, WA · On-site
$200K - $300K/yr
Internship or research experience with LLM inference, ML systems, or model serving * Contributions to open-source inference frameworks (vLLM, SGLang, TensorRT-LLM, etc.) * CUDA / Triton kernel work ...
Member of Technical Staff - Model Optimization and Inference (New Grad)
Seattle, WA · On-site
$200K - $300K/yr
Internship or research experience with LLM inference, ML systems, or model serving * Contributions to open-source inference frameworks (vLLM, SGLang, TensorRT-LLM, etc.) * CUDA / Triton kernel work ...
Machine Learning Engineer (Junior)
New York, NY · On-site
$135K - $150K/yr
Practical experience with deep learning: internships, undergrad or masters' level research projects ... Experience with NVIDIA GPU programming and CUDA * Experience with distributed training frameworks ...
New
Machine Learning Engineer (Junior)
New York, NY · On-site
$135K - $150K/yr
Practical experience with deep learning: internships, undergrad or masters' level research projects ... Experience with NVIDIA GPU programming and CUDA * Experience with distributed training frameworks ...
New
Systems Research Engineer Intern - GPU Programming (Fall 2026)
San Francisco, CA · On-site
$58 - $63/hr
Strong background in GPU programming and parallel computing, such as CUDA and/or Triton ... Our internship dates are September 14th to December 18th. About Together AI Together AI is a ...
Systems Research Engineer Intern - GPU Programming (Fall 2026)
San Francisco, CA · On-site
$58 - $63/hr
Strong background in GPU programming and parallel computing, such as CUDA and/or Triton ... Our internship dates are September 14th to December 18th. About Together AI Together AI is a ...
... CUDA/Triton kernels with custom gradients and tests. • Boost training efficiency and stability ... interns and junior engineers through clear async design docs and code reviews. • Framework ...
... CUDA/Triton kernels with custom gradients and tests. • Boost training efficiency and stability ... interns and junior engineers through clear async design docs and code reviews. • Framework ...
Systems Research Engineer Intern - GPU Programming (Fall 2026)
San Francisco, CA · On-site
$58 - $63/hr
Strong background in GPU programming and parallel computing, such as CUDA and/or Triton ... Our internship dates are September 14th to December 18th. About Together AI Together AI is a ...
Systems Research Engineer Intern - GPU Programming (Fall 2026)
San Francisco, CA · On-site
$58 - $63/hr
Strong background in GPU programming and parallel computing, such as CUDA and/or Triton ... Our internship dates are September 14th to December 18th. About Together AI Together AI is a ...
Systems Research Engineer Intern - GPU Programming (Fall 2026)
San Francisco, CA · On-site
$58 - $63/hr
Strong background in GPU programming and parallel computing, such as CUDA and/or Triton ... Our internship dates are September 14th to December 18th. About Together AI Together AI is a ...
Systems Research Engineer Intern - GPU Programming (Fall 2026)
San Francisco, CA · On-site
$58 - $63/hr
Strong background in GPU programming and parallel computing, such as CUDA and/or Triton ... Our internship dates are September 14th to December 18th. About Together AI Together AI is a ...
Cuda Internship information
See salary details
$2.4K - $2.9K
17% of jobs
$3K is the 25th percentile. Wages below this are outliers.
$2.9K - $3.4K
28% of jobs
$3.4K - $3.9K
3% of jobs
The median wage is $4K / yr.
$3.9K - $4.3K
5% of jobs
$4.3K - $4.8K
0% of jobs
$4.8K - $5.3K
0% of jobs
$5.3K - $5.8K
0% of jobs
$5.8K - $6.3K
0% of jobs
$6.3K - $6.7K
0% of jobs
$6.7K - $7.2K
0% of jobs
$7.4K is the 75th percentile. Wages above this are outliers.
$7.2K - $7.7K
46% of jobs
$2.4K
$5.3K
$7.7K
How much do cuda internship jobs pay per month?
What is a Cuda internship?
A CUDA Internship is a temporary position where interns work with NVIDIA's CUDA parallel computing platform. They typically assist in developing and optimizing GPU-accelerated applications for tasks like machine learning, scientific computing, and gaming. Interns may work on improving algorithms, writing CUDA kernels, or debugging performance issues. This role requires knowledge of C/C++, GPU architectures, and parallel programming concepts. It's ideal for students or recent graduates interested in high-performance computing and GPU programming.
What kind of projects or tasks can I expect to work on during a Cuda internship?
As a CUDA intern, you’ll typically work on tasks related to developing, profiling, or optimizing GPU-accelerated applications and algorithms. Your responsibilities may include writing and testing CUDA kernels, analyzing code performance, and assisting in integrating GPU computing into software projects. You may also collaborate with experienced engineers, learn to use industry-standard tools, and participate in team meetings to discuss technical challenges or progress. This hands-on experience is designed to strengthen your technical skills while giving you insight into real-world GPU development workflows.
What are the key skills and qualifications needed to thrive in the Cuda internship position, and why are they important?
To thrive as a Cuda Intern, you should have a solid background in computer science, strong programming skills (especially in C/C++), and foundational knowledge of parallel computing or GPU architectures. Familiarity with CUDA programming, NVIDIA development tools, and understanding of performance optimization techniques are highly valuable for this role. Strong problem-solving abilities, eagerness to learn, and good teamwork and communication skills will help you excel as an intern. These competencies enable you to contribute effectively to CUDA-based projects and adapt to the fast-paced, innovative environment often found in tech industries.
What cities are hiring for Cuda Internship jobs?
Cities with the most Cuda Internship job openings:
What are the most commonly searched types of Cuda jobs?
The most popular types of Cuda jobs are:
What states have the most Cuda Internship jobs?
States with the most job openings for Cuda Internship jobs include:
What job categories do people searching Cuda Internship jobs look for?
The top searched job categories for Cuda Internship jobs are:

Inference Optimization Intern - Performance Modeling
Sunnyvale, CA • On-site
Internship
Re-posted 27 days ago
Job description
The Institute of Foundation Models is dedicated to advancing the science and engineering of large-scale AI systems. Our researchers and engineers develop cutting-edge foundation models while pushing the limits of high-performance computing and efficient AI inference. By combining deep expertise in machine learning, systems engineering, and hardware optimization, we build scalable AI solutions that drive scientific discovery and real-world impact.
As part of the team, interns work alongside world-class researchers and performance engineers to optimize the execution of large-scale foundation models on next-generation NVIDIA GPU architectures. This internship provides hands-on experience in low-level GPU performance analysis, kernel optimization, and hardware-aware inference acceleration.
Key Responsibilities
This intensive internship offers a unique opportunity to contribute to the development of a simulator and profiling framework for foundation model inference on NVidia GPUs.
Responsibilities include:
- Develop analytical performance models for GPU kernels and inference workloads.
- Build and validate a simulator to estimate theoretical hardware performance limits.
- Compare measured kernel performance against architectural peak throughput.
- Identify performance bottlenecks in compute, memory, communication, and scheduling.
- Analyze GPU execution using NVIDIA Nsight Systems and Nsight Compute.
- Investigate PTX and SASS code generation to understand low-level execution behavior.
- Collaborate with researchers and engineers to optimize inference kernels for transformer-based models.
- Evaluate utilization of Tensor Cores, memory bandwidth, caches, and instruction pipelines.
- Design profiling methodologies for Hopper and Blackwell architectures.
- Document findings and provide actionable recommendations for performance improvements.
Academic Qualifications
Currently pursuing a degree in Computer Science, Computer Engineering, Electrical Engineering, Artificial Intelligence, High-Performance Computing, or a related quantitative discipline.
Preferred Qualifications
- Experience with CUDA programming and GPU kernel development.
- Understanding of NVIDIA GPU architecture and memory hierarchy.
- Familiarity with performance profiling tools such as Nsight Systems and Nsight Compute.
- Knowledge of PTX, SASS, and low-level GPU execution.
- Experience optimizing CUDA kernels for throughput and latency.
- Understanding of roofline analysis, performance modeling, and hardware utilization metrics.
- Experience with deep learning frameworks such as PyTorch or TensorFlow.
- Strong programming skills in C++, CUDA, and Python.
Desired Skills
- Performance engineering mindset.
- Strong analytical and debugging abilities.
- Interest in AI systems, inference optimization, and hardware-software co-design.
- Ability to work independently on research and engineering challenges.
- Excellent written and verbal communication skills.