1

Cluster Manager Jobs (NOW HIRING)

Design and implement a multi-site, highly available Splunk Enterprise deployment including Cluster Manager, License Master, Deployer, * Deployment Server, Monitoring Console, multi-site indexer ...

Posted today

Design and implement a multi-site, highly available Splunk Enterprise deployment including Cluster Manager, License Master, Deployer, * Deployment Server, Monitoring Console, multi-site indexer ...

Posted today

Showing results 21-40

Cluster Manager information

See salary details

$29K

$104.6K

$118K

How much do cluster manager jobs pay per year?

As of Aug 12, 2026, the average yearly pay for cluster manager in the United States is $104,575.00, according to ZipRecruiter salary data. Most workers in this role earn between $114,000.00 and $116,500.00 per year, depending on experience, location, and employer.

What is the difference between Cluster Manager vs Operations Manager?

AspectCluster ManagerOperations Manager
CredentialsBachelor's degree in business, management, or related field; certifications like PMP are commonBachelor's degree in business, management, or related field; certifications like PMP are common
Work EnvironmentOversees multiple locations or units within a region or sectorManages daily operations within a specific department or facility
Industry UsageCommon in retail, healthcare, logistics, and hospitality sectorsWidely used across various industries including manufacturing, retail, and services

While both roles involve management responsibilities, a Cluster Manager oversees multiple sites or units within a region, focusing on strategic coordination, whereas an Operations Manager concentrates on the daily functioning of a specific department or location. The roles often overlap but differ mainly in scope and scale of oversight.

What is the cluster manager?

A cluster manager is a professional responsible for overseeing the operation and management of a group of related servers or systems, often in data centers or cloud environments. They coordinate resources, monitor performance, and ensure system reliability, frequently using tools like Kubernetes or Apache Mesos. Strong organizational skills and technical knowledge of networking and system administration are essential for this role.

What do you need to be a cluster manager?

To become a cluster manager, candidates typically need relevant experience in operations, team leadership, or logistics, along with strong organizational and communication skills. A bachelor's degree in business, management, or a related field is often preferred, and familiarity with industry-specific tools or software can be advantageous.
More about Cluster Manager jobs
What cities are hiring for Cluster Manager jobs? Cities with the most Cluster Manager job openings:
What are the most commonly searched types of Cluster jobs? The most popular types of Cluster jobs are:
What states have the most Cluster Manager jobs? States with the most job openings for Cluster Manager jobs include:
Infographic showing various Cluster Manager job openings in the United States as of August 2026, with employment types broken down into 50% Full Time, and 50% Part Time. Highlights an 100% In-person job distribution, with an average salary of $104,575 per year, or $50.3 per hour.

Principal Software Engineer, Distributed Systems Engineer - DGX Cloud

NVIDIA

Durham, NC • On-site

Full-time

Re-posted 15 hours ago


Nvidia rating

9.6

Company rating: 9.6 out of 10

Based on 17 frontline employees who took The Breakroom Quiz

7th of 243 rated software companies


Job description

Job Summary:
NVIDIA is hiring experienced software engineers with kubernetes experience to help scale up its AI Infrastructure. The role involves working on production systems for large scalable GPU clusters and implementing monitoring capabilities to ensure reliability and performance.
Responsibilities:
• You will be part of an DGX Cloud team responsible for production systems that enable large scalable GPU clusters to be used for a variety of AI workloads. This includes working on custom software related to scheduling GPU resources on kubernetes.
• Implementing monitoring and health management capabilities that enable industry leading reliability, availability, and scalability of GPU assets. You will be harnessing multiple data streams, ranging from GPU hardware diagnostics to cluster and network telemetry.
• Working with teams across NVIDIA to ensure production AI clusters run reliability and consistently with maximum performance. Evaluating system failures and improving services based on a well-defined incident management process.
Qualifications:
Required:
• Significant software engineering experience with kubernetes including cluster operations, operator development, node health monitoring and working with GPU resource scheduling.
• Direct experience in a software engineering role within a highly technical organization with demonstrable impact from your work.
• Software development experience with kubernetes APIs and frameworks not just operating a cluster.
• Highly motivated with strong communication skills, you can work successfully with multi-functional teams, principles, and architects and coordinate effectively across organizational boundaries and geographies.
• 15+ years in similar role and experience on large-scale production systems.
• Experience with common software engineering principles, tools and techniques.
• You possess a BS in Computer Science, Engineering, Physics, Mathematics or a comparable Degree or equivalent experience.
• Technical knowledge, including a systems programming language (Go, Python) and a solid understanding of data structures and algorithms.
Preferred:
• Technical competency in managing and automating large-scale distributed systems independent of cloud providers.
• Advanced hands-on experience and deep understanding of cluster management systems (Kubernetes, Slurm, Bright Cluster Manager).
• Proven operational excellence in maintaining reliable and performant AI infrastructure.
Company:
NVIDIA is a computing platform company operating at the intersection of graphics, HPC, and AI. Founded in 1993, the company is headquartered in Santa Clara, USA, with a team of 10001+ employees. The company is currently Late Stage.

What Nvidia employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Nvidia logo

About Nvidia

Sourced by ZipRecruiter

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.

Industry

Computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

Santa Clara, CA, US

Year founded

1993