UCX
UCX

60 Ucx Jobs Hiring Near You

Senior Deep Learning Communication Architect

Austin, TX · On-site

$128K - $174K/yr

Collaborate with hardware and software teams to craft systems that effectively apply high-speed interconnects (e.g., NVLink, InfiniBand, SPC-X) and communication libraries (e.g., MPI, NCCL, UCX, UCC ...

Senior System Software Engineer

Redmond, WA

$137K - $180K/yr

Work with open source communities to enhance libraries like RAPIDS, CCCL and UCX through technical discussion and code contributions * Provide recommendations and feedback to teams regarding ...

Principal Systems Software Engineer

Champaign, IL · On-site

$135K - $181K/yr

Work with open source communities to enhance libraries like NVIDIA cuDF, CCCL and UCX through technical discussion and code contributions * Collaborate with distributed systems teams to craft ...

Principal Systems Software Engineer

Santa Clara, CA · On-site

$158K - $212K/yr

Work with open source communities to enhance libraries like NVIDIA cuDF, CCCL and UCX through technical discussion and code contributions * Collaborate with distributed systems teams to craft ...

Senior System Software Engineer

Santa Clara, CA · On-site

$143K - $189K/yr

Work with open source communities to enhance libraries like RAPIDS, CCCL and UCX through technical discussion and code contributions * Provide recommendations and feedback to teams regarding ...

UCX for MPI/OpenSHMEM) on GPU clusters. * Participating in and contributing to parallel programming interface specifications like MPI/OpenSHMEM. * Design, implement and maintain system software that ...

Work with open source communities to enhance libraries like RAPIDS, CCCL and UCX through technical discussion and code contributions * Provide recommendations and feedback to teams regarding ...

Showing results 41-60

UCX Jobs Information

What are the most popular states for Ucx jobs?
Infographic showing various job openings at Ucx in the United States as of July 2026, with employment types broken down into 100% Full Time. Highlights an 100% Physical job distribution.
Senior System Software Engineer - GPU Performance

Senior System Software Engineer - GPU Performance

NVIDIA

Santa Clara, CA • On-site

Full-time

Posted 5 days ago


Nvidia rating

9.6

Company rating: 9.6 out of 10

Based on 17 frontline employees who took The Breakroom Quiz

5th of 217 rated software companies


Job description

Job Summary:
NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High Performance Computing and Visualization. The role involves conducting performance characterization and analysis on large multi-GPU and multi-node clusters to influence the roadmap of communication libraries.
Responsibilities:
• Conduct in-depth performance characterization and analysis on large multi-GPU and multi-node clusters.
• Study the interaction of our libraries with all HW (GPU, CPU, Networking) and SW components in the stack
• Evaluate proof-of-concepts, conduct trade-off analysis when multiple solutions are available
• Triage and root-cause performance issues reported by our customers
• Collect a lot of performance data; build tools and infrastructure to visualize and analyze the information
• Collaborate with a very dynamic team across multiple time zones
Qualifications:
Required:
• M.S. (or equivalent experience) or PhD in Computer Science, or related field with relevant performance engineering and HPC experience
• 3+ yrs of experience with parallel programming and at least one communication runtime (MPI, NCCL, UCX, NVSHMEM)
• Experience conducting performance benchmarking and triage on large scale HPC clusters
• Good understanding of computer system architecture, HW-SW interactions and operating systems principles (aka systems software fundamentals)
• Implement micro-benchmarks in C/C++, read and modify the code base when required
• Ability to debug performance issues across the entire HW/SW stack. Proficient in a scripting language, preferably Python
• Familiar with containers, cloud provisioning and scheduling tools (Kubernetes, SLURM, Ansible, Docker)
• Adaptability and passion to learn new areas and tools. Flexibility to work and communicate effectively across different teams and timezones
Preferred:
• Practical experience with Infiniband/Ethernet networks in areas like RDMA, topologies, congestion control
• Experience debugging network issues in large scale deployments
• Familiarity with CUDA programming and/or GPUs
• Experience with Deep Learning Frameworks such PyTorch, TensorFlow
Company:
NVIDIA is a computing platform company operating at the intersection of graphics, HPC, and AI. Founded in 1993, the company is headquartered in Santa Clara, USA, with a team of 10001+ employees. The company is currently Late Stage.

What Nvidia employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Nvidia logo

About Nvidia

Sourced by ZipRecruiter

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.

Industry

Computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

Santa Clara, CA, US

Year founded

1993