1

Infiniband Jobs in California (NOW HIRING)

Required : • 10+ years of experience • Hands-on experience with InfiniBand and Ethernet, including VXLAN and EVPN architectures • Advanced knowledge of EVPN, BGP, OSPF, VXLAN protocols • ...

High-speed networking protocols such as Ethernet , InfiniBand , or RDMA * Routing , switching , and network automation frameworks * Network performance monitoring , telemetry , and diagnostics

Principal ATE Test Engineer

San Jose, CA · On-site

$209K - $240K/yr

Working knowledge of high-speed protocols like PCIe, Ethernet, Infiniband, DDR, NVMe, USB, etc. * Professional attitude with ability to execute on multiple tasks with minimal supervision. * Strong ...

Showing results 41-60

Infiniband information

What is InfiniBand?

Infiniband is a high-speed, low-latency networking technology commonly used in data centers and high-performance computing environments. It is designed to connect servers, storage systems, and network devices, providing much faster data transfer rates than traditional Ethernet. Infiniband supports scalable bandwidth and efficient communication, which makes it ideal for applications requiring rapid data movement, such as scientific simulations and large-scale database transactions. Its architecture also supports remote direct memory access (RDMA), which further reduces latency and CPU overhead.

What are the key skills and qualifications needed to thrive as an InfiniBand network engineer, and why are they important?

To thrive as an InfiniBand Network Engineer, you need a strong background in computer networking, Linux system administration, and high-performance computing (HPC) environments, often supported by a degree in computer science or related field. Familiarity with InfiniBand architecture, experience with tools like OpenFabrics Enterprise Distribution (OFED), and certifications such as CompTIA Network+ are valuable. Strong problem-solving skills, attention to detail, and effective communication are crucial soft skills for this role. These abilities are essential for ensuring efficient, reliable InfiniBand network performance in complex HPC or data center environments.

What are the typical responsibilities of an InfiniBand network engineer in a data center environment?

InfiniBand network engineers are primarily responsible for designing, deploying, and maintaining high-performance InfiniBand fabrics that connect servers and storage systems in data centers, especially in HPC (High-Performance Computing) environments. Their daily tasks include monitoring network performance, troubleshooting connectivity or latency issues, and performing firmware and driver updates on InfiniBand switches and host adapters. They also collaborate closely with system administrators and application teams to optimize throughput and ensure reliable, low-latency communication. Additionally, InfiniBand engineers often participate in capacity planning and help scale the network infrastructure to meet growing computational demands.

What is the difference between Infiniband vs Ethernet Network Engineer?

AspectInfinibandEthernet Network Engineer
Required CredentialsNetworking certifications, Cisco, Cisco CCNA, CCNPNetworking certifications, Cisco, CCNA, CCNP
Work EnvironmentData centers, high-performance computing environmentsCorporate networks, data centers, enterprise environments
Industry UsageHigh-performance computing, research institutionsBusiness, telecommunications, enterprise IT
Common Search/ComparisonYesYes

Infiniband and Ethernet Network Engineers both work with network infrastructure, but Infiniband specializes in high-speed, low-latency connections used in data centers and HPC environments. Ethernet Network Engineers focus on standard Ethernet networks used across various industries. While their certifications and skills overlap, their work environments and applications differ significantly.

What job categories do people searching Infiniband jobs in California look for? The top searched job categories for Infiniband jobs in California are:
What cities in California are hiring for Infiniband jobs? Cities in California with the most Infiniband job openings:
Infographic showing various Infiniband job openings in California as of August 2026, with employment types broken down into 87% Full Time, 11% Part Time, and 2% Contract. Highlights an 86% Physical, 3% Hybrid, and 11% Remote job distribution.

Senior Software Engineer - Together Cloud Infrastructure

Together AI

San Francisco, CA • On-site

$127K - $173K/yr

Full-time

Re-posted 2 days ago


Job description

Job Summary:
Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle. As a Senior AI Infrastructure Engineer, you will be responsible for building a highly available cloud infrastructure that supports cutting-edge ML hardware and serves both internal products and external customers.
Responsibilities:
• Design, build, and maintain performant, secure, and highly-available backend services/operators that run in our data centers and automate hardware management, such as Infiniband partitioning, in-DC parallel storage provisioning, and VM provisioning.
• Design and build out the IaaS software layer for a new GB200 data center with thousands of GPUs.
• Work on a global multi-exabyte high-performance object store, serving massive datasets for pretraining.
• Build advanced observability stacks for our customers with automated node lifecycle management for fault-tolerant distributed pretraining.
• Perform architecture and research work for decentralized AI workloads
• Work on the core, open-source Together AI platform
• Create services, tools, and developer documentation
• Create testing frameworks for robustness and fault-tolerance
Qualifications:
Required:
• 5+ years of professional software development experience and proficiency in at least one backend programming language (Golang desired)
• 5+ years experience writing high-performance, well-tested, production quality code
• Demonstrated experience with building and operating high-performance and/or globally distributed micro-service architectures across one or more cloud providers (AWS, Azure, GCP)
• Excellent communication skills – able to write clear design docs and work effectively with both technical and non-technical team members
• Strong systems knowledge across compute, networking, and storage, including concurrency, memory management, performant I/O, and scale
• Experience with infrastructure automation tools (Terraform, Ansible), monitoring/observability stacks (Prometheus, Grafana), and CI/CD pipelines (GitHub Actions, ArgoCD)
Preferred:
• Deep experience with Kubernetes internals a big plus, such as implementing non-trivial Kubernetes operators, device/storage/network plugins, custom schedulers, or patches thereon or Kubernetes itself
• Deep experience with VMs/hypervisors a big plus, such as QEMU/KVM, cloud-hypervisor, VFIO, virtio, PCIE passthrough, Kubevirt, SR-IOV
• Deep experience with DC networking tech + solutions a big plus, such as VLAN, VXLAN, VPN, VPC, OVS/OVN
• Experience with Cluster API or similar a big plus
• Experience working on high-performance compute, networking, and/or storage a big plus
• Experience virtualizing GPUs and/or Infiniband a big plus
• Experience building IaaS or PaaS systems at scale a plus
• Experience with DPUs/SmartNICs a plus
• GPU programming, NCCL, CUDA knowledge a plus
Company:
Together AI provides a cloud platform for developing, training, fine-tuning, and deploying generative AI models. Founded in 2022, the company is headquartered in San Francisco, USA, with a team of 201-500 employees. The company is currently Growth Stage.