Staff Engineer, Inference Optimizations
Boston, MA · Remote
$191K - $239K/yr
Expert-level Triton or CUDA. If you've contributed to the Triton compiler or wrote custom CUDA ... This is a remote role JR: 2026-7625 #LI-Remote
Boston, MA · Remote
$191K - $239K/yr
Expert-level Triton or CUDA. If you've contributed to the Triton compiler or wrote custom CUDA ... This is a remote role JR: 2026-7625 #LI-Remote
Boston, MA · Remote
$191K - $239K/yr
Expert-level Triton or CUDA. If you've contributed to the Triton compiler or wrote custom CUDA ... This is a remote role JR: 2026-7625 #LI-Remote
Boston, MA · On-site +1
Design and maintain high-performance GPU kernels in Triton or CUDA for state-of-the-art ML ... remote.
Boston, MA · On-site +1
Design and maintain high-performance GPU kernels in Triton or CUDA for state-of-the-art ML ... remote.
Boston, MA · On-site +1
$146K - $225K/yr
Strong programming skills in C++ and/or CUDA programming * This opportunity can support remote work within the United States, with occasional travel. Working Conditions While performing the duties of ...
Boston, MA · On-site +1
$146K - $225K/yr
Strong programming skills in C++ and/or CUDA programming * This opportunity can support remote work within the United States, with occasional travel. Working Conditions While performing the duties of ...
Boston, MA · On-site +1
$146K - $225K/yr
Strong programming skills in C++ and/or CUDA programming * This opportunity can support remote work within the United States, with occasional travel. Working Conditions While performing the duties of ...
Quick apply
Boston, MA · On-site +1
$146K - $225K/yr
Strong programming skills in C++ and/or CUDA programming * This opportunity can support remote work within the United States, with occasional travel. Working Conditions While performing the duties of ...
Boston, MA · On-site +1
$146K - $225K/yr
Strong programming skills in C++ and/or CUDA programming * This opportunity can support remote work within the United States, with occasional travel. Working Conditions While performing the duties of ...
Boston, MA · On-site +1
$146K - $225K/yr
Strong programming skills in C++ and/or CUDA programming * This opportunity can support remote work within the United States, with occasional travel. Working Conditions While performing the duties of ...
C++/CUDA a plus * Experience with distributed ML training frameworks (Megatron-LM, TorchTitan ... Experience training MoE architectures Location San Francisco, CA or Cambridge, MA (Remote, Hybrid ...
C++/CUDA a plus * Experience with distributed ML training frameworks (Megatron-LM, TorchTitan ... Experience training MoE architectures Location San Francisco, CA or Cambridge, MA (Remote, Hybrid ...
Burlington, MA · Remote
$150K - $185K/yr
THIS IS NOT A FULLY REMOTE POSITION. Required * Bachelor's degree in Computer Science, Electrical ... Ability to collaborate effectively across software, firmware, DevOps, data science, and hardware ...
Quick apply
Burlington, MA · Remote
$150K - $185K/yr
THIS IS NOT A FULLY REMOTE POSITION. Required * Bachelor's degree in Computer Science, Electrical ... Ability to collaborate effectively across software, firmware, DevOps, data science, and hardware ...
Cambridge, MA · On-site +1
$135K - $177K/yr
Implement and optimize kernels using CUDA, OpenMP, OpenCL, or accelerator-specific programming ... Hybrid / primarily remote within approved payroll states Qualifications Basic Qualifications are ...
Cambridge, MA · On-site +1
$135K - $177K/yr
Implement and optimize kernels using CUDA, OpenMP, OpenCL, or accelerator-specific programming ... Hybrid / primarily remote within approved payroll states Qualifications Basic Qualifications are ...
Boston, MA · On-site +1
$144K - $192K/yr
Design and maintain high-performance GPU kernels in Triton or CUDA for state-of-the-art ML ... be fully remote. The salary range for this role is an estimate based on a wide range of ...
Quick apply
Boston, MA · On-site +1
$144K - $192K/yr
Design and maintain high-performance GPU kernels in Triton or CUDA for state-of-the-art ML ... be fully remote. The salary range for this role is an estimate based on a wide range of ...
Collaborate with ML and software engineering colleagues to deploy and operationalize models ... Comfort working with modern ML infrastructure (e.g., Docker, CUDA, Kubernetes, experiment tracking ...
Collaborate with ML and software engineering colleagues to deploy and operationalize models ... Comfort working with modern ML infrastructure (e.g., Docker, CUDA, Kubernetes, experiment tracking ...
$90.7K - $95.8K
17% of jobs
$97.5K is the 25th percentile. Wages below this are outliers.
$95.8K - $101K
26% of jobs
The median wage is $102.9K / yr.
$101K - $106.1K
20% of jobs
$106.1K - $111.3K
4% of jobs
$111.3K - $116.4K
5% of jobs
$119.6K is the 75th percentile. Wages above this are outliers.
$116.4K - $121.5K
4% of jobs
$121.5K - $126.7K
5% of jobs
$126.7K - $131.8K
4% of jobs
$131.8K - $136.9K
4% of jobs
$136.9K - $142.1K
5% of jobs
$142.1K - $147.2K
4% of jobs
$90.7K
$111.4K
$147.2K
| Aspect | Remote Cuda Developer | Remote Machine Learning Engineer |
|---|---|---|
| Required Credentials | CUDA programming certifications, computer science degree | Machine learning certifications, data science background |
| Work Environment | Software development, GPU optimization | Model development, data analysis |
| Industry Usage | High-performance computing, gaming, AI | AI, data science, predictive modeling |
Remote Cuda Developers focus on GPU programming and optimization using CUDA, primarily in high-performance computing and AI applications. Remote Machine Learning Engineers develop and deploy machine learning models, often utilizing GPU resources but with a broader focus on data and algorithms. While both roles may involve GPU expertise, Cuda Developers specialize in low-level programming, whereas Machine Learning Engineers work on model development and deployment.
$191K - $239K/yr
Other
Posted 3 days ago
DigitalOcean is seeking a Senior Engineer 2 to play a key technical role in our AI Inference Optimization team. DigitalOcean aims to be the Inference Cloud of choice for digitally native companies and you will help ensure we can offer the industry-leading performance for our inference services. You will be responsible for the architectural decisions that maximize throughput and minimize latency for the world's most advanced large models. As an IC leader, you will act as a force multiplier for the engineering organization, solving the most complex bottlenecks in memory bandwidth and compute utilization while guiding the technical roadmap for our high-performance inference fleet.
What You'll Do:*This is a remote role
JR: 2026-7625
#LI-Remote
Sourced by ZipRecruiter
Software development
501 - 1,000 Employees
New York, NY, US
2012