Lab Administrator
Columbia, MD · On-site
Experience with HPC/AI infrastructure (InfiniBand, RDMA, GPU clusters) * Experience with configuration management (Ansible, Puppet)
Columbia, MD · On-site
Experience with HPC/AI infrastructure (InfiniBand, RDMA, GPU clusters) * Experience with configuration management (Ansible, Puppet)
Columbia, MD · On-site
Experience with HPC/AI infrastructure (InfiniBand, RDMA, GPU clusters) * Experience with configuration management (Ansible, Puppet)
Santa Clara, CA · On-site
$143K/yr
Experience with Linux systems, distributed GPU infrastructure, RDMA/RoCE networking, or communication libraries such as NCCL/RCCL is highly valued. KEY RESPONSIBILITIES: * Full-stack GPU profiling:
Santa Clara, CA · On-site
$143K/yr
Experience with Linux systems, distributed GPU infrastructure, RDMA/RoCE networking, or communication libraries such as NCCL/RCCL is highly valued. KEY RESPONSIBILITIES: * Full-stack GPU profiling:
Santa Clara, CA · On-site
$123K - $169K/yr
Experience with Linux systems, distributed GPU infrastructure, RDMA/RoCE networking, or communication libraries such as NCCL/RCCL is highly valued. KEY RESPONSIBILITIES: * Full-stack GPU profiling:
Santa Clara, CA · On-site
$123K - $169K/yr
Experience with Linux systems, distributed GPU infrastructure, RDMA/RoCE networking, or communication libraries such as NCCL/RCCL is highly valued. KEY RESPONSIBILITIES: * Full-stack GPU profiling:
San Jose, CA · On-site
$164K - $202K/yr
Role: GPU Software Engineer/GPU Architect Location: San Jose, CA (Remote/Hybrid) Duration ... RDMA, RoCE, InfiniBand, or Infinity Fabric * Distributed inference/training or HPC experience
San Jose, CA · On-site
$164K - $202K/yr
Role: GPU Software Engineer/GPU Architect Location: San Jose, CA (Remote/Hybrid) Duration ... RDMA, RoCE, InfiniBand, or Infinity Fabric * Distributed inference/training or HPC experience
Manhattan, NY · On-site
$160 - $260/hr
RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request ... Own Baseten's GPU fabric observability and root-cause analysis architecture. * Build telemetry ...
Manhattan, NY · On-site
$160 - $260/hr
RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request ... Own Baseten's GPU fabric observability and root-cause analysis architecture. * Build telemetry ...
San Francisco, CA · On-site
$190 - $270/hr
RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request ... Own Baseten's GPU fabric observability and root‑cause analysis architecture. * Build telemetry ...
San Francisco, CA · On-site
$190 - $270/hr
RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request ... Own Baseten's GPU fabric observability and root‑cause analysis architecture. * Build telemetry ...
San Francisco, CA · On-site
$160 - $260/hr
RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request ... Own Baseten's GPU fabric observability and root-cause analysis architecture. * Build telemetry ...
San Francisco, CA · On-site
$160 - $260/hr
RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request ... Own Baseten's GPU fabric observability and root-cause analysis architecture. * Build telemetry ...
New York, NY · On-site
$200K - $380K/yr
RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request ... Own Baseten's GPU fabric observability and root-cause analysis architecture. * Build telemetry ...
New York, NY · On-site
$200K - $380K/yr
RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request ... Own Baseten's GPU fabric observability and root-cause analysis architecture. * Build telemetry ...
... as RDMA, GPU Direct Storage, and distributed filesystems protocols such as NFS or FUSE to optimize storage performance and efficiency. • Lead efforts to improve the reliability, durability ...
... as RDMA, GPU Direct Storage, and distributed filesystems protocols such as NFS or FUSE to optimize storage performance and efficiency. • Lead efforts to improve the reliability, durability ...
... as RDMA, GPU Direct Storage, and distributed filesystems protocols such as NFS or FUSE to optimize storage performance and efficiency. • Lead efforts to improve the reliability, durability ...
... as RDMA, GPU Direct Storage, and distributed filesystems protocols such as NFS or FUSE to optimize storage performance and efficiency. • Lead efforts to improve the reliability, durability ...
Alibaba Cloud is seeking a skilled RDMA Ops Engineer to optimize and maintain high-performance ... and GPU-aware scheduling • Background in Computing system optimization (NVIDIA collective ...
Alibaba Cloud is seeking a skilled RDMA Ops Engineer to optimize and maintain high-performance ... and GPU-aware scheduling • Background in Computing system optimization (NVIDIA collective ...
Collaborate with other teams to architect RDMA-capable hardware and define transport layer optimizations for GPU-based large scale AI workload deployments. * Use and modify system models, perform ...
Collaborate with other teams to architect RDMA-capable hardware and define transport layer optimizations for GPU-based large scale AI workload deployments. * Use and modify system models, perform ...
... RDMA/InfiniBand optimization experience • Contributions to GPU libraries or frameworks • Low-level debugging skills (PTX/SASS reading) Company : Genmo is an artificial intelligence creative ...
... RDMA/InfiniBand optimization experience • Contributions to GPU libraries or frameworks • Low-level debugging skills (PTX/SASS reading) Company : Genmo is an artificial intelligence creative ...
Collaborate with other teams to architect RDMA-capable hardware and define transport layer optimizations for GPU-based large scale AI workload deployments. * Use and modify system models, perform ...
Collaborate with other teams to architect RDMA-capable hardware and define transport layer optimizations for GPU-based large scale AI workload deployments. * Use and modify system models, perform ...
Foundation knowledge of high-performance networking technologies, Ethernet, RoCE/RDMA, GPU accelerated computing, generative AI, LLM training and inference, Network SerDes, in-network compute, PCIe ...
Foundation knowledge of high-performance networking technologies, Ethernet, RoCE/RDMA, GPU accelerated computing, generative AI, LLM training and inference, Network SerDes, in-network compute, PCIe ...
San Francisco, CA · On-site
$150K - $180K/yr
RDMA Fabric Selection: Design and justify InfiniBand (NDR/XDR, Quantum-class) vs. lossless Ethernet ... Kubernetes networking for GPU serving (CNI, SR-IOV, Multus). * Expert certifications (CCIE/JNCIE ...
San Francisco, CA · On-site
$150K - $180K/yr
RDMA Fabric Selection: Design and justify InfiniBand (NDR/XDR, Quantum-class) vs. lossless Ethernet ... Kubernetes networking for GPU serving (CNI, SR-IOV, Multus). * Expert certifications (CCIE/JNCIE ...
San Francisco, CA · On-site
$170 - $230/hr
RDMA Fabric Selection: Design and justify InfiniBand (NDR/XDR, Quantum-class) vs. lossless Ethernet ... Kubernetes networking for GPU serving (CNI, SR-IOV, Multus). * Expert certifications (CCIE/JNCIE ...
San Francisco, CA · On-site
$170 - $230/hr
RDMA Fabric Selection: Design and justify InfiniBand (NDR/XDR, Quantum-class) vs. lossless Ethernet ... Kubernetes networking for GPU serving (CNI, SR-IOV, Multus). * Expert certifications (CCIE/JNCIE ...
Oversee high-performance GPU interconnects (InfiniBand/Ethernet fabrics), optimizing spine-leaf architectures, RDMA, network telemetry, and overall performance.High-Performance Storage Integration:
Oversee high-performance GPU interconnects (InfiniBand/Ethernet fabrics), optimizing spine-leaf architectures, RDMA, network telemetry, and overall performance.High-Performance Storage Integration:
Austin, TX · On-site +1
$180K - $320K/yr
Configure and optimize high-performance host networking stacks, including RDMA, SR-IOV, RoCEv2, and ... Implement and manage sophisticated GPU slicing technologies (MIG, vGPU) to enable efficient multi ...
Austin, TX · On-site +1
$180K - $320K/yr
Configure and optimize high-performance host networking stacks, including RDMA, SR-IOV, RoCEv2, and ... Implement and manage sophisticated GPU slicing technologies (MIG, vGPU) to enable efficient multi ...
$180 - $240/hr
Configure and optimize high-performance host networking stacks, including RDMA, SR-IOV, RoCEv2, and ... Implement and manage sophisticated GPU slicing technologies (MIG, vGPU) to enable efficient multi ...
$180 - $240/hr
Configure and optimize high-performance host networking stacks, including RDMA, SR-IOV, RoCEv2, and ... Implement and manage sophisticated GPU slicing technologies (MIG, vGPU) to enable efficient multi ...
$42.5K - $54.5K
0% of jobs
$54.5K - $66.6K
0% of jobs
$66.6K - $78.6K
2% of jobs
$78.6K - $90.7K
7% of jobs
$90.7K - $102.7K
13% of jobs
$104.9K is the 25th percentile. Wages below this are outliers.
$102.7K - $114.8K
18% of jobs
The median wage is $121.5K / yr.
$114.8K - $126.8K
19% of jobs
$126.8K - $138.9K
16% of jobs
$140K is the 75th percentile. Wages above this are outliers.
$138.9K - $150.9K
11% of jobs
$150.9K - $163K
9% of jobs
$163K - $175K
5% of jobs
$42.5K
$123.8K
$175K
| Aspect | Rdma Gpu | Network Engineer |
|---|---|---|
| Required Credentials | Computer science or related degree, certifications in GPU computing or high-performance networking | Networking certifications (CCNA, CCNP), degree in computer science or related field |
| Work Environment | Data centers, high-performance computing labs, research facilities | Corporate offices, data centers, telecommunication environments |
| Industry Usage | AI, machine learning, scientific computing, data analytics | IT infrastructure, network design, security, and maintenance |
Rdma Gpu specialists focus on optimizing GPU performance and high-speed data transfer using RDMA technology, primarily in computing and research environments. Network Engineers design, implement, and maintain network systems. While both roles involve high-tech infrastructure, Rdma Gpu roles are more specialized in GPU and high-performance data transfer, whereas Network Engineers focus on network connectivity and security.
Cities with the most Rdma Gpu job openings:
States with the most job openings for Rdma Gpu jobs include:
The top searched job categories for Rdma Gpu jobs are:

Other
This job post has expired 3 days ago. Applications are no longer accepted.
Responsible for the setup, maintenance, security, and day-to-day operation of a technical/research lab environment (servers, storage systems, networking equipment, and lab-issued workstations), ensuring high availability, performance, and compliance with organizational policies.
Key Responsibilities