Staff Engineer, Inference Optimizations
Denver, CO · Remote
$191K - $239K/yr
Excellent system design skills, particularly related to low-level GPU programming - optimization ... This is a remote role JR: 2026-7625 #LI-Remote
Denver, CO · Remote
$191K - $239K/yr
Excellent system design skills, particularly related to low-level GPU programming - optimization ... This is a remote role JR: 2026-7625 #LI-Remote
Denver, CO · Remote
$191K - $239K/yr
Excellent system design skills, particularly related to low-level GPU programming - optimization ... This is a remote role JR: 2026-7625 #LI-Remote
Denver, CO · On-site +1
$100K - $135K/yr
Remote USA - In Tandem Compensation: $100,000 - $135,000 / year Description At In Tandem, we build ... Run the inference serving layer on our own GPU hardware: choose and tune the serving stack (vLLM ...
Denver, CO · On-site +1
$100K - $135K/yr
Remote USA - In Tandem Compensation: $100,000 - $135,000 / year Description At In Tandem, we build ... Run the inference serving layer on our own GPU hardware: choose and tune the serving stack (vLLM ...
... GPU acceleration, memory optimization) * Knowledge of DoD or Intelligence Community mission systems, especially related to remote sensing or space-based sensors * Experience transitioning algorithms ...
... GPU acceleration, memory optimization) * Knowledge of DoD or Intelligence Community mission systems, especially related to remote sensing or space-based sensors * Experience transitioning algorithms ...
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Experience implementing algorithms on the GPU in Python or C++ using CUDA and other CUDA libraries
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Experience implementing algorithms on the GPU in Python or C++ using CUDA and other CUDA libraries
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Experience implementing algorithms on the GPU in Python or C++ using CUDA and other CUDA libraries
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Experience implementing algorithms on the GPU in Python or C++ using CUDA and other CUDA libraries
Englewood, CO · Remote
$70.75 - $91/hr
Remote ( Travel may be required; candidate must be flexible across U.S. time zones) Duration: 6+ ... Collaborate across departments including software development, DevOps, QA, and training * Educate ...
Quick apply
Englewood, CO · Remote
$70.75 - $91/hr
Remote ( Travel may be required; candidate must be flexible across U.S. time zones) Duration: 6+ ... Collaborate across departments including software development, DevOps, QA, and training * Educate ...
Location This is a fully remote opportunity based in US East or US West, with travel as needed ... No engineering lift. The problem it solves is real and measurable. In DoiT's survey of 500 finance ...
Quick apply
Location This is a fully remote opportunity based in US East or US West, with travel as needed ... No engineering lift. The problem it solves is real and measurable. In DoiT's survey of 500 finance ...
$34K - $39.8K
5% of jobs
$39.8K - $45.7K
10% of jobs
$45.7K - $51.5K
7% of jobs
$52.6K is the 25th percentile. Wages below this are outliers.
$51.5K - $57.4K
15% of jobs
$57.4K - $63.2K
7% of jobs
The median wage is $65.3K / yr.
$63.2K - $69.1K
15% of jobs
$69.1K - $74.9K
11% of jobs
$79.3K is the 75th percentile. Wages above this are outliers.
$74.9K - $80.8K
6% of jobs
$80.8K - $86.6K
14% of jobs
$86.6K - $92.4K
7% of jobs
$92.4K - $98.3K
2% of jobs
$34K
$66.9K
$98.3K

$191K - $239K/yr
Full-time
Posted 16 days ago
DigitalOcean is seeking a Senior Engineer 2 to play a key technical role in our AI Inference Optimization team. DigitalOcean aims to be the Inference Cloud of choice for digitally native companies and you will help ensure we can offer the industry-leading performance for our inference services. You will be responsible for the architectural decisions that maximize throughput and minimize latency for the world's most advanced large models. As an IC leader, you will act as a force multiplier for the engineering organization, solving the most complex bottlenecks in memory bandwidth and compute utilization while guiding the technical roadmap for our high-performance inference fleet.
What You'll Do:*This is a remote role
JR: 2026-7625
#LI-Remote
Sourced by ZipRecruiter
Software development
501 - 1,000 Employees
New York, NY, US
2012