1

Datacenter Operations Jobs (NOW HIRING)

$135K - $150K/yr

Oversees enterprise datacenter operations and hybrid-cloud integration. Responsibilities: * Manage servers, SAN, and virtualization environments. * Lead backup/disaster recovery planning and testing.

Datacenter technician

Abernathy, TX · On-site

$23 - $25/hr

Role- DC technician Location - 1414 Farm to Market Road 54, Abernathy, TX 79311 L1 Data Center Technician Exp 3+ years About the Team The Datacenter Operation team supports the company's fast growth ...

Data Center Lead

Monterey, CA · On-site

$135K - $150K/yr

Oversees enterprise datacenter operations and hybrid-cloud integration. Responsibilities: * Manage servers, SAN, and virtualization environments. * Lead backup/disaster recovery planning and testing.

Showing results 41-60

Datacenter Operations information

See salary details

$52K

$128.5K

$200K

How much do datacenter operations jobs pay per year?

As of Aug 10, 2026, the average yearly pay for datacenter operations in the United States is $128,526.00, according to ZipRecruiter salary data. Most workers in this role earn between $94,000.00 and $163,500.00 per year, depending on experience, location, and employer.

What is the difference between Datacenter Operations vs Data Center Technician?

AspectDatacenter OperationsData Center Technician
CertificationsCompTIA Server+, Cisco CCNA, Data Center certificationsCompTIA Server+, Cisco CCNA, Data Center certifications
Work EnvironmentData centers, server rooms, 24/7 operationsData centers, server rooms, hardware troubleshooting
Job FocusMonitoring, managing infrastructure, ensuring uptimeInstalling, maintaining, troubleshooting hardware
Employer & Industry UsageData center providers, IT departmentsData center providers, IT support teams

Both roles operate within data centers and require similar certifications. However, Datacenter Operations focuses on managing and monitoring infrastructure to ensure continuous uptime, while Data Center Technicians primarily handle hardware installation and troubleshooting. Understanding these differences helps in choosing the right career path or job search focus.

What is datacenter operations?

Datacenter Operations refer to the processes, activities, and staff responsible for managing and maintaining the technical infrastructure within a datacenter. This includes overseeing servers, storage, networking hardware, power, cooling systems, and security to ensure continuous uptime and optimal performance. Datacenter operations teams are also tasked with monitoring systems, troubleshooting issues, performing regular maintenance, and implementing disaster recovery plans. Their work is crucial to ensure that organizational data and IT services are available, secure, and reliable.

What are some common challenges faced in a datacenter operations role, and how can they be effectively managed?

Professionals in Datacenter Operations often encounter challenges such as maintaining high availability, managing physical and cybersecurity risks, and responding to hardware failures or outages quickly. Effective management involves following strict protocols, proactive monitoring, and collaborating closely with IT, network, and facilities teams to ensure seamless operations. Additionally, staying up-to-date with evolving technologies and adopting automation tools can help address these challenges and improve overall efficiency.

What are the key skills and qualifications needed to thrive in datacenter operations?

To thrive in Datacenter Operations, you need a solid understanding of IT infrastructure, hardware troubleshooting, and network management, typically supported by relevant certifications such as CompTIA Server+, Network+, or Cisco CCNA. Familiarity with data center management tools, monitoring systems, and ticketing platforms is commonly required. Strong problem-solving skills, attention to detail, and effective communication are essential soft skills for this role. These abilities ensure system reliability, efficient issue resolution, and seamless collaboration in maintaining critical business operations.
More about Datacenter Operations jobs
What cities are hiring for Datacenter Operations jobs? Cities with the most Datacenter Operations job openings:
What are the most commonly searched types of Datacenter Operations jobs? The most popular types of Datacenter Operations jobs are:
What states have the most Datacenter Operations jobs? States with the most job openings for Datacenter Operations jobs include:
Infographic showing various Datacenter Operations job openings in the United States as of August 2026, with employment types broken down into 85% Full Time, 12% Part Time, 1% Temporary, and 2% Contract. Highlights an 94% Physical, 2% Hybrid, and 4% Remote job distribution, with an average salary of $128,526 per year, or $61.8 per hour.

Staff Engineer, Datacenter Server Lifecycle

Anthropic

San Francisco, CA • On-site

Full-time

Re-posted 13 days ago


Job description

About the role

Anthropic is investing $50 billion in American computing infrastructure, including datacenters custom-built for our workloads, and this role sits at the heart of that effort. As a Staff Engineer on the Datacenter Server Lifecycle team, you will own the end-to-end operational journey of every machine in our datacenters - from initial provisioning and deployment, steady-state operation, maintenance, and repair. 

This is greenfield work: you will help define the services, tooling, standards and processes that govern how we operate critical  hardware at scale, powering frontier model development and serving. You will have the opportunity to define AI-native workflows for datacenter operations and to help drive innovations on efficiency, performance, and reliability.

A distinguishing aspect of this role is its deep intersection with security. The machines in our datacenter handle some of the most sensitive workloads in AI - training frontier models and serving millions of users interacting with Claude. Ensuring that every machine in the fleet is trusted, attested, and operating with a verified chain of integrity from the hardware up is a core part of the job, not an afterthought. You will partner closely with our Infrastructure Security team to define and enforce trusted compute standards across the lifecycle, from secure provisioning through end-of-life handling.

Key responsibilities
  • Build automation to support datacenter fleets at scale.
  • Define and own the end-to-end system lifecycle strategy - from provisioning and deployment through operation, maintenance, refresh, and decommissioning - and maintain automation and operational procedures for common lifecycle events (e.g., hardware failures, firmware upgrades, fleet rotations).
  • Partner closely with Infrastructure Security to design and enforce trusted compute standards across the server lifecycle.
  • Work closely with our Networking team to ensure end-to-end connectivity across all sites.
  • Build and maintain tooling to track machine health, configuration, and operational status across the full datacenter fleet.
Minimum qualifications
  • Hands-on experience with server hardware, including rack deployment, cabling, troubleshooting, and understanding failure modes at scale.
  • End-to-end understanding of hardware lifecycle management: asset tracking, provisioning workflows, maintenance scheduling, and decommissioning practices.
  • Proficiency in at least one programming language (e.g., Python, Rust, Go, or Java).
  • Working knowledge of modern cloud infrastructure, including Kubernetes and large scale cloud providers (e.g. AWS, Azure, GCP).
  • Ability to communicate clearly and build consensus with a wide range of stakeholders.
  • Comfort navigating ambiguity and making progress on complex, cross-functional problems.
  • Willingness to travel occasionally to datacenter sites across North America.
Preferred qualifications
  • 8+ years of experience in datacenter infrastructure management, or a closely related discipline.
  • Hands-on experience with GPU or AI accelerator hardware (e.g., NVIDIA A100/H100, Google TPUs, or AWS Trainium) and an understanding of their operational demands.
  • Familiarity with modern provisioning and OS tooling such as LinuxBoot and NixOS
  • Experience building or contributing to datacenter automation or fleet management platforms.
  • Experience building and deploying server operating system distributions across large server fleets.
  • Background in large-scale capacity planning and hardware refresh strategy, ideally at a hyperscaler or large cloud provider.
  • Experience with trusted compute and hardware security concepts such as secure boot, TPM, hardware attestation, and firmware verification - or a strong desire to develop deep expertise in this area.