1

Cluster Manager Jobs in Raleigh, NC (NOW HIRING)

Azure DevOps Architect

Durham, NC · Remote

$61.25 - $80/hr

... cluster management for large-scale AKS environments Required Experience: - 10+ years of experience in DevOps architecture and cloud platform engineering. - Demonstrated enterprise-scale delivery ...

New

They are responsible for storage provisioning, cluster management, data protection, security, and performance tuning across both SAN FC, iSCSI and NAS NFS, SMBCIFS environments to ensure high ...

Azure DevOps Architect

Durham, NC · On-site

$80 - $90/hr

... cluster management for large-scale AKS environments Required Experience: - 10+ years of experience in DevOps architecture and cloud platform engineering. - Demonstrated enterprise-scale delivery ...

New

Manage Hadoop stack support run book. Responsible for cluster availability Support development and production deployments Support disaster recovery and business continuity practice Required Skills ...

The primary purpose of this position is to ensure the uniform, high-quality delivery of both pre-award and post-award services across their assigned cluster of academic departments. The Hub Manager ...

next page

Showing results 1-20

Cluster Manager information

See Raleigh, NC salary details

$28.2K

$101.7K

$114.7K

How much do cluster manager jobs pay per year?

As of Aug 12, 2026, the average yearly pay for cluster manager in Raleigh, NC is $101,656.00, according to ZipRecruiter salary data. Most workers in this role earn between $110,800.00 and $113,200.00 per year, depending on experience, location, and employer.

What is the difference between Cluster Manager vs Operations Manager?

AspectCluster ManagerOperations Manager
CredentialsBachelor's degree in business, management, or related field; certifications like PMP are commonBachelor's degree in business, management, or related field; certifications like PMP are common
Work EnvironmentOversees multiple locations or units within a region or sectorManages daily operations within a specific department or facility
Industry UsageCommon in retail, healthcare, logistics, and hospitality sectorsWidely used across various industries including manufacturing, retail, and services

While both roles involve management responsibilities, a Cluster Manager oversees multiple sites or units within a region, focusing on strategic coordination, whereas an Operations Manager concentrates on the daily functioning of a specific department or location. The roles often overlap but differ mainly in scope and scale of oversight.

What is the cluster manager?

A cluster manager is a professional responsible for overseeing the operation and management of a group of related servers or systems, often in data centers or cloud environments. They coordinate resources, monitor performance, and ensure system reliability, frequently using tools like Kubernetes or Apache Mesos. Strong organizational skills and technical knowledge of networking and system administration are essential for this role.

What do you need to be a cluster manager?

To become a cluster manager, candidates typically need relevant experience in operations, team leadership, or logistics, along with strong organizational and communication skills. A bachelor's degree in business, management, or a related field is often preferred, and familiarity with industry-specific tools or software can be advantageous.
What are popular job titles related to Cluster Manager jobs in Raleigh, NC? For Cluster Manager jobs in Raleigh, NC, the most frequently searched job titles are:
What job categories do people searching Cluster Manager jobs in Raleigh, NC look for? The top searched job categories for Cluster Manager jobs in Raleigh, NC are:
Infographic showing various Cluster Manager job openings in Raleigh, NC as of June 2026, with employment types broken down into 93% Full Time, 5% Part Time, 1% Temporary, and 1% Contract. Highlights an 89% Physical, 1% Hybrid, and 10% Remote job distribution, with an average salary of $101,650 per year, or $48.9 per hour.

Principal Software Engineer, Distributed Systems Engineer - DGX Cloud

NVIDIA

Durham, NC • On-site

Full-time

Re-posted 21 hours ago


Nvidia rating

9.6

Company rating: 9.6 out of 10

Based on 17 frontline employees who took The Breakroom Quiz

7th of 243 rated software companies


Job description

Job Summary:
NVIDIA is hiring experienced software engineers with kubernetes experience to help scale up its AI Infrastructure. The role involves working on production systems for large scalable GPU clusters and implementing monitoring capabilities to ensure reliability and performance.
Responsibilities:
• You will be part of an DGX Cloud team responsible for production systems that enable large scalable GPU clusters to be used for a variety of AI workloads. This includes working on custom software related to scheduling GPU resources on kubernetes.
• Implementing monitoring and health management capabilities that enable industry leading reliability, availability, and scalability of GPU assets. You will be harnessing multiple data streams, ranging from GPU hardware diagnostics to cluster and network telemetry.
• Working with teams across NVIDIA to ensure production AI clusters run reliability and consistently with maximum performance. Evaluating system failures and improving services based on a well-defined incident management process.
Qualifications:
Required:
• Significant software engineering experience with kubernetes including cluster operations, operator development, node health monitoring and working with GPU resource scheduling.
• Direct experience in a software engineering role within a highly technical organization with demonstrable impact from your work.
• Software development experience with kubernetes APIs and frameworks not just operating a cluster.
• Highly motivated with strong communication skills, you can work successfully with multi-functional teams, principles, and architects and coordinate effectively across organizational boundaries and geographies.
• 15+ years in similar role and experience on large-scale production systems.
• Experience with common software engineering principles, tools and techniques.
• You possess a BS in Computer Science, Engineering, Physics, Mathematics or a comparable Degree or equivalent experience.
• Technical knowledge, including a systems programming language (Go, Python) and a solid understanding of data structures and algorithms.
Preferred:
• Technical competency in managing and automating large-scale distributed systems independent of cloud providers.
• Advanced hands-on experience and deep understanding of cluster management systems (Kubernetes, Slurm, Bright Cluster Manager).
• Proven operational excellence in maintaining reliable and performant AI infrastructure.
Company:
NVIDIA is a computing platform company operating at the intersection of graphics, HPC, and AI. Founded in 1993, the company is headquartered in Santa Clara, USA, with a team of 10001+ employees. The company is currently Late Stage.

What Nvidia employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Nvidia logo

About Nvidia

Sourced by ZipRecruiter

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.

Industry

Computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

Santa Clara, CA, US

Year founded

1993