1

Datacenter Operations Jobs in Texas (NOW HIRING)

Responsibilities Datacenter Operations * Complies with Data Center business unit policies, procedures, and deadlines with guidance from experienced technicians and/or direct-line management.

Work with our Datacenter Operations Engineers to maintain and operate the fleet of AI systems at peak performance. * Drive corrective actions for systems that are not operating correctly, working ...

Work with our Datacenter Operations Engineers to maintain and operate the fleet of AI systems at peak performance. * Drive corrective actions for systems that are not operating correctly, working ...

Work with our Datacenter Operations Engineers to maintain and operate the fleet of AI systems at peak performance. * Drive corrective actions for systems that are not operating correctly, working ...

Work with our Datacenter Operations Engineers to maintain and operate the fleet of AI systems at peak performance. * Drive corrective actions for systems that are not operating correctly, working ...

The Datacenter Technician will ensure the continuous operation and security of critical infrastructure while performing maintenance and overseeing day-to-day facility operations. Responsibilities ...

This individual will be responsible for ensuring 24x7 continuous operation and security of critical infrastructure including all routine and preventive maintenance activities. As a Datacenter ...

We are seeking a Datacenter Manager to lead data center operations, infrastructure management, and engineering initiatives. * The ideal candidate will oversee a team of engineers, ensure reliability ...

Showing results 41-60

Datacenter Operations information

See Texas salary details

$48.4K

$119.7K

$186.3K

How much do datacenter operations jobs pay per year?

As of Aug 9, 2026, the average yearly pay for datacenter operations in Texas is $119,742.00, according to ZipRecruiter salary data. Most workers in this role earn between $87,600.00 and $152,300.00 per year, depending on experience, location, and employer.

What is the difference between Datacenter Operations vs Data Center Technician?

AspectDatacenter OperationsData Center Technician
CertificationsCompTIA Server+, Cisco CCNA, Data Center certificationsCompTIA Server+, Cisco CCNA, Data Center certifications
Work EnvironmentData centers, server rooms, 24/7 operationsData centers, server rooms, hardware troubleshooting
Job FocusMonitoring, managing infrastructure, ensuring uptimeInstalling, maintaining, troubleshooting hardware
Employer & Industry UsageData center providers, IT departmentsData center providers, IT support teams

Both roles operate within data centers and require similar certifications. However, Datacenter Operations focuses on managing and monitoring infrastructure to ensure continuous uptime, while Data Center Technicians primarily handle hardware installation and troubleshooting. Understanding these differences helps in choosing the right career path or job search focus.

What is datacenter operations?

Datacenter Operations refer to the processes, activities, and staff responsible for managing and maintaining the technical infrastructure within a datacenter. This includes overseeing servers, storage, networking hardware, power, cooling systems, and security to ensure continuous uptime and optimal performance. Datacenter operations teams are also tasked with monitoring systems, troubleshooting issues, performing regular maintenance, and implementing disaster recovery plans. Their work is crucial to ensure that organizational data and IT services are available, secure, and reliable.

What are some common challenges faced in a datacenter operations role, and how can they be effectively managed?

Professionals in Datacenter Operations often encounter challenges such as maintaining high availability, managing physical and cybersecurity risks, and responding to hardware failures or outages quickly. Effective management involves following strict protocols, proactive monitoring, and collaborating closely with IT, network, and facilities teams to ensure seamless operations. Additionally, staying up-to-date with evolving technologies and adopting automation tools can help address these challenges and improve overall efficiency.

What are the key skills and qualifications needed to thrive in datacenter operations?

To thrive in Datacenter Operations, you need a solid understanding of IT infrastructure, hardware troubleshooting, and network management, typically supported by relevant certifications such as CompTIA Server+, Network+, or Cisco CCNA. Familiarity with data center management tools, monitoring systems, and ticketing platforms is commonly required. Strong problem-solving skills, attention to detail, and effective communication are essential soft skills for this role. These abilities ensure system reliability, efficient issue resolution, and seamless collaboration in maintaining critical business operations.
What are the most commonly searched types of Datacenter Operations jobs in Texas? The most popular types of Datacenter Operations jobs in Texas are:
What job categories do people searching Datacenter Operations jobs in Texas look for? The top searched job categories for Datacenter Operations jobs in Texas are:
Infographic showing various Datacenter Operations job openings in Texas as of August 2026, with employment types broken down into 85% Full Time, 12% Part Time, 1% Temporary, and 2% Contract. Highlights an 94% Physical, 2% Hybrid, and 4% Remote job distribution, with an average salary of $119,742 per year, or $57.6 per hour.

Head of AI Inference & MLOps

Deeter Analytics

Austin, TX • On-site

Other

Posted 4 days ago


Job description

Overview

Location: Austin, Texas area / On-site preferred

Project: 7MW Phase I AI Datacenter -> 50MW Campus Expansion

Reports to: Founders / Executive Team

About the Project

We are building a high-density AI datacenter campus outside Austin, Texas, beginning with approximately 7MW of NVIDIA GB300 NVL72 infrastructure and scaling to 50MW+. The initial deployment is designed around real-time inference, reasoning, and high-value AI serving workloads, with a focus on monetizing capacity in live markets rather than simply leasing powered space.

This is not a traditional datacenter operations role.

We are hiring the person who will make the racks make money.

This leader will own the strategy and execution required to turn rack-scale GPU infrastructure into a profitable inference business: selecting the right models, runtimes, orchestration stack, routing layer, pricing strategy, customer segments, and marketplace relationships to maximize revenue, uptime, and utilization.

The right candidate understands that raw compute is not the business. Monetized tokens, latency-adjusted utilization, and gross margin are the business.

The Role

We need a senior operator-builder who can sit at the intersection of:

  • AI infrastructure

  • inference performance engineering

  • model serving and routing

  • marketplace monetization

  • customer / partner integration

  • revenue optimization

You will design and run the inference platform that determines how our GB300 NVL72 racks are monetized in the real-time market. That may include direct enterprise workloads, marketplace distribution, API-based reselling, model hosting, fine-tuned/private deployments, and emerging inference channels.

You should know what makes money on modern inference hardware, what does not, and why.

You should be able to answer questions like:

  • Which open-weight and commercial-compatible models should run on this hardware first?

  • How should workloads be split between premium low-latency serving, bulk throughput, reserved capacity, and experimental capacity?

  • Should we route through third-party marketplaces, sell directly, or do both?

  • What software stack gives us the best performance per watt, per GPU, and per dollar of capex?

  • How do we maximize realized revenue rather than theoretical benchmark performance?

  • How do we scale from a 7MW launch to a repeatable 50MW AI factory operating model?

What You’ll Own

  • Build and lead the inference monetization strategy for our first 7MW deployment and expansion to 50MW

  • Define the technical and commercial operating model for turning GB300 NVL72 racks into revenue-producing assets

  • Evaluate and implement the model serving stack, scheduling layer, inference engine, observability stack, and API platform

  • Select and optimize the mix of workloads across:

    • real-time inference

    • reasoning workloads

    • premium low-latency API traffic

    • batch / overflow workloads

    • dedicated enterprise deployments

    • private/fine-tuned model hosting

  • Identify the best go-to-market channels for capacity monetization, including direct sales and marketplace/API distribution partners

  • Develop strategy for integration with platforms such as OpenRouter-style aggregation, OpenAI-compatible endpoints, and other inference distribution channels where appropriate. OpenRouter provides a unified API and provider aggregation layer, while Inference.net offers an OpenAI-compatible API experience around model access and deployment, making both relevant examples of the ecosystem this role would evaluate. (OpenRouter)

  • Own benchmarking methodology based on actual profit and production metrics, not vanity metrics

  • Drive workload placement decisions based on revenue per rack, revenue per GPU-hour, revenue per MW, latency targets, and customer value

  • Partner with datacenter engineering, networking, and facilities teams to ensure the physical plant supports the intended software monetization strategy

  • Build pricing, SLAs, utilization strategy, and customer segmentation framework

  • Create dashboards and control systems for:

    • utilization

    • queue health

    • latency

    • token throughput

    • margin by workload

    • failure rate

    • realized revenue by cluster / rack / model / customer

  • Lead decisions around multi-tenant vs single-tenant deployments, reserved vs on-demand capacity, and when to prioritize direct contracts over marketplace traffic

  • Build and manage the team required to scale this function over time

What Success Looks Like

In the first 3–6 months, you will:

  • Stand up a production inference platform for our initial GB300 NVL72 deployment

  • Recommend the highest-value initial workloads and monetization channels

  • Launch a repeatable commercialization strategy for rack capacity

  • Establish a clear performance and revenue measurement framework

  • Identify where we should sell capacity: direct, through marketplaces, via strategic partners, or through a hybrid approach

  • Turn the first cluster into a measurable cash-generating operation

In the first 12 months, you will:

  • Build the operating playbook for scaling from 7MW to 50MW

  • Increase utilization without destroying margins or SLA quality

  • Improve realized revenue per rack through model, routing, pricing, and customer mix optimization

  • Establish the company as a serious real-time inference operator, not just a GPU owner

Required Experience
  • Significant experience in production AI/LLM inference, MLOps, model serving, or AI infrastructure monetization

  • Proven experience running or scaling GPU-backed inference systems in production

  • Strong understanding of modern inference runtimes, serving frameworks, and optimization techniques

  • Experience with one or more of:

    • vLLM

    • TensorRT-LLM

    • SGLang

    • Ray Serve

    • Triton Inference Server

    • Kubernetes-based GPU orchestration

    • custom routing / scheduler layers

  • Experience optimizing for real-world production metrics such as throughput, latency, GPU utilization, availability, and cost efficiency

  • Strong understanding of LLM inference economics, including tradeoffs among model size, quantization, latency, throughput, memory footprint, and customer willingness to pay

  • Experience building or managing API-based AI platforms or inference products

  • Ability to translate infrastructure capability into a pricing and product strategy

  • Experience working with enterprise customers, developer platforms, or AI marketplaces

  • Strong technical judgment on model selection, infrastructure topology, and commercialization strategy

Preferred Experience
  • Experience monetizing large-scale NVIDIA GPU infrastructure

  • Experience with rack-scale or cluster-scale inference environments

  • Background in both technical operations and business strategy

  • Familiarity with AI inference aggregators, routing platforms, and model marketplaces

  • Experience designing multi-tenant GPU systems with strong isolation and predictable performance

  • Experience with advanced observability, token-level metering, cost accounting, and SLA enforcement

  • Familiarity with reasoning-model workloads, agentic inference, multimodal inference, and future high-density AI factory architectures

  • Experience supporting OpenAI-compatible APIs and enterprise private deployments

What Makes Someone Great in This Role
  • You know the difference between “high benchmark performance” and “high realized revenue”

  • You understand that some workloads are great for utilization but terrible for margin

  • You can spot when a shiny model is commercially useless

  • You know how to tune systems for the workloads customers will actually pay for

  • You are opinionated about the stack, but flexible about the business model

  • You can go deep technically and still think like an owner

Compensation

Competitive salary, bonus, and equity participation tied to the scale, importance, and revenue generated from the role.

#J-18808-Ljbffr