Background in virtualization, multi-tenancy, confidential computing, or cloud GPU provisioning. * Contributions to open-source kernel, driver, or runtime projects. ACADEMIC CREDENTIALS: * Bachelor ...
Background in virtualization, multi-tenancy, confidential computing, or cloud GPU provisioning. * Contributions to open-source kernel, driver, or runtime projects. ACADEMIC CREDENTIALS: * Bachelor ...
Productionize models from the research team, spanning containerization, inference optimization, and deployment to edge devices and cloud GPU infrastructure. * Build and own offline and online ...
Productionize models from the research team, spanning containerization, inference optimization, and deployment to edge devices and cloud GPU infrastructure. * Build and own offline and online ...
Principal System Software Architect, AI/GPU Platforms
Austin, TX · On-site
$178K/yr
Background in virtualization, multi-tenancy, confidential computing, or cloud GPU provisioning. * Contributions to open-source kernel, driver, or runtime projects. ACADEMIC CREDENTIALS: * Bachelor ...
Principal System Software Architect, AI/GPU Platforms
Austin, TX · On-site
$178K/yr
Background in virtualization, multi-tenancy, confidential computing, or cloud GPU provisioning. * Contributions to open-source kernel, driver, or runtime projects. ACADEMIC CREDENTIALS: * Bachelor ...
$200 - $250/hr
Productionize models from the research team, spanning containerization, inference optimization, and deployment to edge devices and cloud GPU infrastructure. * Build and own offline and online ...
$200 - $250/hr
Productionize models from the research team, spanning containerization, inference optimization, and deployment to edge devices and cloud GPU infrastructure. * Build and own offline and online ...
Performance Engineer
Palo Alto, CA · On-site
... cloud environments • Optimize latency, throughput, memory usage, batching, scheduling, routing, and GPU utilization • Investigate performance regressions in real customer environments • Work ...
Performance Engineer
Palo Alto, CA · On-site
... cloud environments • Optimize latency, throughput, memory usage, batching, scheduling, routing, and GPU utilization • Investigate performance regressions in real customer environments • Work ...
Performance Engineer
Palo Alto, CA · On-site
... cloud environments • Optimize latency, throughput, memory usage, batching, scheduling, routing, and GPU utilization • Investigate performance regressions in real customer environments • Work ...
Performance Engineer
Palo Alto, CA · On-site
... cloud environments • Optimize latency, throughput, memory usage, batching, scheduling, routing, and GPU utilization • Investigate performance regressions in real customer environments • Work ...
Systems Engineer
Redmond, WA · On-site
Responsibilities : • Design, build, and optimize systems infrastructure spanning edge devices to cloud GPU clusters for robotics workloads. • Develop and maintain low-latency, high-throughput ...
Systems Engineer
Redmond, WA · On-site
Responsibilities : • Design, build, and optimize systems infrastructure spanning edge devices to cloud GPU clusters for robotics workloads. • Develop and maintain low-latency, high-throughput ...
Delivery Director, Capacity Programs
San Francisco, CA · On-site
$200 - $250/hr
Drive neo cloud delivery programs -- manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
New
Delivery Director, Capacity Programs
San Francisco, CA · On-site
$200 - $250/hr
Drive neo cloud delivery programs -- manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
New
Delivery Director, Capacity Programs
San Francisco, CA · On-site
$200 - $250/hr
Drive neo cloud delivery programs -- manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
Delivery Director, Capacity Programs
San Francisco, CA · On-site
$200 - $250/hr
Drive neo cloud delivery programs -- manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
Managing Director, EdgeUno Compute
Miami, FL · On-site
$150 - $200/hr
Role Overview We are looking for a Managing Director, EdgeUno Compute, to lead and scale our GPU, bare metal, and cloud infrastructure business across Latin America and the Americas. This role ...
Managing Director, EdgeUno Compute
Miami, FL · On-site
$150 - $200/hr
Role Overview We are looking for a Managing Director, EdgeUno Compute, to lead and scale our GPU, bare metal, and cloud infrastructure business across Latin America and the Americas. This role ...
A leading AI infrastructure company is seeking a Senior Solution Architect to design innovative GPU cloud and AI solutions. In this role, you'll engage with enterprise and hyperscaler customers ...
A leading AI infrastructure company is seeking a Senior Solution Architect to design innovative GPU cloud and AI solutions. In this role, you'll engage with enterprise and hyperscaler customers ...
Delivery Director, Capacity Programs
San Francisco, CA · On-site
$200 - $250/hr
Drive neo cloud delivery programs -- manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
New
Delivery Director, Capacity Programs
San Francisco, CA · On-site
$200 - $250/hr
Drive neo cloud delivery programs -- manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
New
$80 - $100/hr
With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal ...
$80 - $100/hr
With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal ...
$200 - $250/hr
Drive neo cloud delivery programs -- manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
New
$200 - $250/hr
Drive neo cloud delivery programs -- manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
New
Delivery Director, Capacity Programs
San Francisco, CA · On-site
$265K - $285K/yr
Drive neo cloud delivery programs - manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
Delivery Director, Capacity Programs
San Francisco, CA · On-site
$265K - $285K/yr
Drive neo cloud delivery programs - manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
$200 - $250/hr
Drive neo cloud delivery programs -- manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
New
$200 - $250/hr
Drive neo cloud delivery programs -- manage capacity delivery from GPU cloud and neo cloud partners (e.g., colocation/bare-metal/GPU cloud providers), including contract milestones, capacity ramps ...
New
Role Overview We are looking for a Managing Director, EdgeUno Compute, to lead and scale our GPU, bare metal, and cloud infrastructure business across Latin America and the Americas. This role ...
Role Overview We are looking for a Managing Director, EdgeUno Compute, to lead and scale our GPU, bare metal, and cloud infrastructure business across Latin America and the Americas. This role ...
Staff AI/ML Infrastructure Engineer
$145K - $160K/yr
With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal ...
Staff AI/ML Infrastructure Engineer
$145K - $160K/yr
With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal ...
Senior Technical Product Manager, Observability
$130K - $165K/yr
With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal ...
Senior Technical Product Manager, Observability
$130K - $165K/yr
With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal ...
Software Engineer, Core Cloud Engineering
$80K - $95K/yr
With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal ...
Software Engineer, Core Cloud Engineering
$80K - $95K/yr
With 33 global cloud data center locations, Vultr is trusted by hundreds of thousands of active customers across 185 countries for its flexible, scalable, global Cloud Compute, Cloud GPU, Bare Metal ...
Cloud Gpu information
See salary details
$10.82 - $17.55
4% of jobs
$17.55 - $24.28
2% of jobs
$24.28 - $31.01
1% of jobs
$31.01 - $37.74
1% of jobs
$37.74 - $44.47
3% of jobs
$44.47 - $51.20
6% of jobs
$53.97 is the 25th percentile. Wages below this are outliers.
$51.20 - $57.93
18% of jobs
The median wage is $62.64 / hr.
$57.93 - $64.66
21% of jobs
$64.66 - $71.39
15% of jobs
$72.98 is the 75th percentile. Wages above this are outliers.
$71.39 - $78.12
18% of jobs
$78.12 - $84.86
11% of jobs
$10
$61
$84
How much do cloud gpu jobs pay per hour?
What is a cloud GPU?
What skills and qualifications are needed to work with cloud GPUs?
What are the main challenges faced when managing cloud GPU resources in a production environment?
What is the difference between Cloud Gpu vs Data Scientist?
| Aspect | Cloud Gpu | Data Scientist |
|---|---|---|
| Required Credentials | Knowledge of cloud platforms, GPU computing, and basic programming | Degree in data science, statistics, or related field; often Python or R skills |
| Work Environment | Cloud-based infrastructure, hardware management, and GPU resources | Data analysis, modeling, and visualization in office or remote settings |
| Industry Usage | Tech, AI, machine learning, and high-performance computing | Business, finance, healthcare, and research sectors |
While Cloud Gpu specialists focus on managing GPU resources in cloud environments for high-performance tasks, Data Scientists analyze data to extract insights and build models. Both roles often collaborate in AI projects but differ in technical focus and daily tasks.
What other helpful pages are available for Cloud Gpu?
Other pages related to Cloud Gpu:

Principal System Software Architect, AI/GPU Platforms
Austin, TX • Hybrid
Full-time
Posted 20 days ago
Advanced Micro Devices rating
8.6
Based on 13 frontline employees who took The Breakroom Quiz
28th of 162 rated electronics manufacturers
Job description
WHAT YOU DO AT AMD CHANGES EVERYTHING
At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career.
THE ROLE
You will join the system software architecture team behind AMD Instinct™ accelerators, the GPUs powering some of the world's largest AI and HPC deployments. This role is focused on next-generation, rack-scale AI platforms in the MI400 class spanning the GPU, the node, and the scale-up/scale-out fabric that binds thousands of accelerators into a single training and inference system.
THE PERSONAs a Systems Software Architect, you sit at the intersection of silicon, firmware, driver, runtime, and framework. You define how the software stack exposes and orchestrates the hardware so that AMD's largest customers can extract maximum performance, reliability, and utilization from their infrastructure.
KEY RESPONSIBILITIES- Own the end-to-end system software architecture for one or more MI400-class subsystems for example GPU memory management, scheduling and queuing, RAS and serviceability, virtualization/partitioning (SR-IOV), or the scale-up/scale-out interconnect software model.
- Drive architecture across the stack: kernel-mode driver (amdgpu/KFD), user-mode runtime (ROCr/HSA), firmware interfaces, and the ROCm software platform, ensuring the layers compose cleanly and perform.
- Partner with silicon and SoC architects during pre-silicon definition to shape hardware/software interfaces, programming models, and register/firmware contracts before tape-out.
- Define the software strategy for multi-GPU and rack-scale topologies, including Infinity Fabric / UALink-style interconnect, collective communication (RCCL), memory coherence, and address translation across the platform.
- Establish architecture for reliability, availability, and serviceability at scale, error detection, containment, telemetry, recovery, and graceful degradation across large clusters.
- Set direction on performance: identify bottlenecks in the launch path, memory subsystem, and communication path, and define the software mechanisms to close them.
- Produce architecture specifications, reference designs, and design reviews that align firmware, driver, runtime, and framework teams onto a shared plan.
- Act as a technical anchor across AMD and with strategic hyperscale and AI customers — translating their workload requirements into architectural direction and representing AMD in deep technical engagements.
PREFERRED EXPERIENCE
- Linux Memory Management and Heterogeneous Memory Management (HMM).
- GPU / DRM driver development.
- Cache coherence and memory consistency protocols.
- GPU Networking technologies including including scale up transport, NVLink, UALink, RDMA, and peer-direct.
- Scale-up and scale-out networking, and communication collectives (e.g., RCCL/NCCL), MPI, or SHMEM.
- Data-center fabrics such as Infinity Fabric, UALink, Ultra Ethernet, InfiniBand, and RoCE, with topology-aware software.
DESIRABLE EXPERIENCE
- Direct experience with the ROCm stack, AMD Instinct, CUDA, or comparable GPU compute ecosystems.
- Hands-on experience developing or optimizing GPU compute kernels (HIP, CUDA, Triton, or assembly-level tuning).
- Experience with deep learning frameworks like PyTorch and TensorFlow in particular including framework integration, custom operators, and performance tuning on GPU backends.
- Familiarity with the broader ML framework and compiler ecosystem (PyTorch, JAX, TensorFlow, ONNX, MLIR/compiler stacks).
- Hands-on experience with GPU compute technologies such as OpenCL and Vulkan.
- Familiarity with AI/ML training and inference workloads (transformers, large-scale distributed training, KV-cache and memory pressure, inference serving).
- Background in virtualization, multi-tenancy, confidential computing, or cloud GPU provisioning.
- Contributions to open-source kernel, driver, or runtime projects.
ACADEMIC CREDENTIALS:
- Bachelor’s or Master’s in Electrical Engineer, Computer Engineering, Computer Science, or a closely related field
LOCATION:
Austin, TX
Santa Clara, CA
This role is not eligible for visa sponsorship.
#LI-BW2
#LI-HYBRID
Benefits offered are described: AMD benefits at a glance.
AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.
AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.
This posting is for an existing vacancy.
Qualifications:Benefits offered are described: AMD benefits at a glance.
AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.
AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.
This posting is for an existing vacancy.
Education:UNAVAILABLEEmployment Type: FULL_TIMEWhat Advanced Micro Devices employees say
Pay
Benefits
Hours and flexibility
Workplace
Get the full story on Breakroom
About Advanced Micro Devices (AMD)
Sourced by ZipRecruiter
Industry
Computer and electronic product manufacturing and manufacturing
Company size
5,001 - 10,000 Employees
Headquarters location
Sunnyvale, CA, US