1

Cuda Programming Jobs in New Rochelle, NY (NOW HIRING)

Write and optimize code using CUDA, PTX assembly, and architecture-specific techniques * Apply advanced performance optimization methods such as memory coalescing, warp-level programming, tensor core ...

Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten's Model ... Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA ...

Posted today

Experience with NVIDIA GPU programming and CUDA * Experience with distributed training frameworks, such as DeepSpeed, FSDL, Ray * Experience with inference frameworks like vLLM * Experience with ...

New

Experience with NVIDIA GPU programming and CUDA * Experience with distributed training frameworks, such as DeepSpeed, FSDL, Ray * Experience with inference frameworks like vLLM * Experience with ...

New

Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten's Model ... Profile and optimize TensorRT‑LLM kernels, analyze CUDA kernel performance, implement custom CUDA ...

Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Baseten ... Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA ...

Showing results 41-60

Cuda Programming information

See New Rochelle, NY salary details

$28

$55

$84

How much do cuda programming jobs pay per hour?

As of Aug 21, 2026, the average hourly pay for cuda programming in New Rochelle, NY is $55.94, according to ZipRecruiter salary data. Most workers in this role earn between $45.29 and $65.29 per hour, depending on experience, location, and employer.

What is the difference between Cuda Programming vs GPU Developer?

AspectCuda ProgrammingGPU Developer
Required CredentialsKnowledge of CUDA, C/C++, parallel computingKnowledge of GPU architecture, CUDA, OpenCL, C/C++
Work EnvironmentHigh-performance computing, scientific research, AIGraphics, gaming, scientific visualization, AI
Industry UsageTech companies, research labs, AI firmsGaming, entertainment, tech, research

While Cuda Programming focuses specifically on writing code using NVIDIA's CUDA platform for parallel processing, GPU Developers have a broader role that includes designing, optimizing, and implementing GPU-based solutions across various platforms and technologies. Both roles require knowledge of GPU architecture and programming languages like C/C++, but GPU Developers often work on a wider range of applications beyond CUDA-specific projects.

Are CUDA programmers in demand?

CUDA programmers are in high demand due to the growing need for high-performance computing in fields like artificial intelligence, scientific research, and data processing. Skills in parallel programming, GPU architecture, and CUDA toolkit are highly valued, and job opportunities are expected to grow as industries adopt GPU acceleration for complex tasks.

What does a CUDA programming developer do?

A CUDA programming developer writes software that leverages NVIDIA's CUDA platform to perform parallel processing on GPUs, optimizing computational tasks such as scientific simulations, machine learning, and image processing. They typically work with C++ and CUDA-specific libraries, debugging and optimizing code for high performance in environments that require intensive data processing.

What cities near New Rochelle, NY are hiring for Cuda Programming jobs?

Cities near New Rochelle, NY with the most Cuda Programming job openings:

Infographic showing various Cuda Programming job openings in New Rochelle, NY as of August 2026, with employment types broken down into 1% Internship, 79% Full Time, 14% Part Time, 1% Temporary, 4% Contract, and 1% Nights. Highlights an 88% Physical, 3% Hybrid, and 9% Remote job distribution, with an average salary of $116,347 per year, or $55.9 per hour.

Campus AI Research Engineer - Deep Learning (Intern)

Jump Trading

Manhattan, NY • On-site

$300K/yr

Other

Re-posted 13 days ago


Job description

Campus Ai Research Engineer - Deep Learning (Intern)

Jump Trading Group is committed to world class research. We empower exceptional talents in Mathematics, Physics, and Computer Science to seek scientific boundaries, push through them, and apply cutting edge research to global financial markets. Our culture is unique. Constant innovation requires fearlessness, creativity, intellectual honesty, and a relentless competitive streak. We believe in winning together and unlocking unique individual talent by incenting collaboration and mutual respect. At Jump, research outcomes drive more than superior risk adjusted returns. We design, develop, and deploy technologies that change our world, fund start-ups across industries, and partner with leading global research organizations and universities to solve problems.

Our trading teams are each comprised of a dynamic group of traders, quantitative researchers, and engineers who work together to examine the global markets, seeking to understand the complexities of various traded products and exchanges. They leverage their impeccable statistical analysis and data mining skills, using the results of their research to make forecasts and develop profitable predictive trading models.

We are seeking research scientists with a demonstrated ability to apply machine learning to achieve state-of-the-art capabilities in complex and challenging domains. The ideal person for this role will be capable of implementing an open-ended research project from concept to production and continuously improving model design, tools, and infrastructure. Potential projects may target any area of the quantitative research and monetization process. We believe that successful research efforts require a fluid mix of skills including AI/ML expertise, engineering pragmatism, statistics, and market intuition.

What You'll Do:

  • Apply state-of-the-art techniques to complex and challenging domains.
  • Work closely with researchers and quants to build flexible and reusable frameworks for financial ML.
  • Optimize training pipelines to make the best use of our HPC resources.
  • Integrate ML models into production systems where latency matters.
  • Work across a mix of programming languages: C / C++ / Python / CUDA and other low-level GPU languages.
  • Build large-scale ML systems that are observable, performant, and flexible. Help improve productivity by reducing the iteration cycle time on research.
  • Other duties as assigned or needed.

Skills You'll Need:

  • Strong publication record at ICML, ICLR, AAAI, NeurIPS, UAI, KDD, or equivalent and/or contributions to open-source AI research
  • Strong general ML background with exposure to modern deep learning techniques and/or language modeling architectures (e.g. transformers, SSMs)
  • Solid development skills in Python and/or C++
  • Familiarity with ML libraries/frameworks such as PyTorch, JAX, and/or TensorFlow
  • Intellectual curiosity, versatility, and originality combined with a pragmatic outlook
  • Ability to thrive in a collaborative, team-oriented environment
  • Ability to reason through quantitative problems and communicate effectively with trading researchers
  • Reliable and predictable availability

Bonus Points:

  • Experience with HPC and distributed large model training
  • Experience with GPU performance optimization (CUDA or ROCm)
  • Experience with end-to-end model development
  • Strong opinions on best practices in ML research, tooling, and/or infrastructure

INTERNATIONAL STUDENTS are encouraged to apply. We accept students eligible for CPT/OPT and we sponsor work visas for full-time positions.

The estimated base salary for this role (annualized) is $300,000 per year.