1

Infrastructure Operations Engineer Jobs (NOW HIRING)

AI Infrastructure Operations Engineer

Columbus, OH · Hybrid

$103K - $136K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

AI Infrastructure Operations Engineer

Tampa, FL · Hybrid

$101K - $133K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

AI Infrastructure Operations Engineer

Redmond, WA · Hybrid

$120K - $157K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

$93K - $123K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

AI Infrastructure Operations Engineer

Nashville, TN · Hybrid

$103K - $136K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

AI Infrastructure Operations Engineer

Boston, MA · Hybrid

$116K - $153K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

AI Infrastructure Operations Engineer

Dallas, TX · Hybrid

$106K - $139K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

AI Infrastructure Operations Engineer

Carmel, NY · Hybrid

$108K - $142K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

AI Infrastructure Operations Engineer

Miami, FL · Hybrid

$102K - $134K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

AI Infrastructure Operations Engineer

Raleigh, NC · Hybrid

$104K - $137K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

AI Infrastructure Operations Engineer

New York, NY · Hybrid

$117K - $154K/yr

Experience building AI infrastructure automation and operations tools, AgenticOps practices that ... engineering, change management, and performance validation for AI training, inference, HPC, and ...

Showing results 41-60

Infrastructure Operations Engineer information

See salary details

$46.5K

$127.1K

$182K

How much do infrastructure operations engineer jobs pay per year?

As of Sep 7, 2026, the average yearly pay for infrastructure operations engineer in the United States is $127,066.00, according to ZipRecruiter salary data. Most workers in this role earn between $107,500.00 and $141,000.00 per year, depending on experience, location, and employer.

What does an infrastructure operations engineer do?

An Infrastructure Operations Engineer is responsible for maintaining and supporting an organization's IT infrastructure, such as servers, networks, and cloud platforms. Their duties include monitoring system performance, troubleshooting issues, managing deployments, and ensuring high availability and security of infrastructure components. They often work with other IT teams to implement upgrades, automate processes, and optimize system reliability. This role is critical in keeping business operations running smoothly and efficiently.

What are some common challenges faced by infrastructure operations engineers, and how can they be addressed?

Infrastructure Operations Engineers often encounter challenges such as managing complex systems, ensuring high availability, and responding quickly to incidents or outages. Balancing proactive maintenance with reactive troubleshooting is key, as is staying up to date with evolving technologies. Building strong collaboration with development, security, and support teams helps streamline workflows and minimizes downtime. Adopting automation tools and clear documentation can also alleviate repetitive tasks and improve incident response times.

What are the key skills and qualifications needed to thrive as an infrastructure operations engineer, and why are they important?

To thrive as an Infrastructure Operations Engineer, you need solid knowledge of networking, system administration, cloud platforms, and scripting, usually backed by a degree in computer science or a related field. Familiarity with tools like Linux/Windows servers, AWS/Azure, monitoring systems, and certifications such as CompTIA Network+ or AWS Certified SysOps Administrator are highly valuable. Outstanding problem-solving skills, attention to detail, and effective communication are essential soft skills in this role. These abilities ensure reliable infrastructure performance, minimize downtime, and support seamless business operations.

What is the difference between Infrastructure Operations Engineer vs Network Engineer?

AspectInfrastructure Operations EngineerNetwork Engineer
CertificationsCompTIA Network+, Cisco CCNA, VMware certificationsCompTIA Network+, Cisco CCNA, Juniper JNCIA
Work EnvironmentData centers, cloud environments, enterprise IT infrastructureNetwork infrastructure, routers, switches, firewalls
Industry UsageIT service providers, large enterprises, cloud providersTelecommunications, enterprise IT, internet service providers

While both roles involve network and infrastructure knowledge, Infrastructure Operations Engineers focus on maintaining overall IT infrastructure, including servers and cloud systems, whereas Network Engineers specialize in designing, implementing, and managing network hardware and connectivity. Understanding these distinctions helps in choosing the right career path or job search focus.

Infographic showing various Infrastructure Operations Engineer job openings in the United States as of August 2026, with employment types broken down into 88% Full Time, 11% Part Time, and 1% Contract. Highlights an 94% Physical, 2% Hybrid, and 4% Remote job distribution, with an average salary of $127,066 per year, or $61.1 per hour.

AI Infrastructure Operations Engineer

Accenture

Los Angeles, CA • Hybrid

$115K - $151K/yr

Full-time

Medical, Dental, Vision, Life, Retirement, PTO

Posted 5 days ago


Accenture Federal Services rating

8.7

Company rating: 8.7 out of 10

Based on 20 frontline employees who took The Breakroom Quiz

51st of 500 rated business services


Job description

Accenture is a global professional services company with leading capabilities in digital, cloud and security. Combining unmatched experience and specialized skills across more than 40 industries, we offer Strategy and, Interactive, Technology, and Operations services, all powered by the world's largest network of Advanced Technology and Intelligent Operations centers. Our 738,000 people deliver on the promise of technology and human ingenuity every day, serving clients in more than 120 countries. We embrace the power of change to create value and shared success for our clients, people, shareholders, partners, and communities. Visit us atwww.accenture.com.

The Global AI Infrastructure team enables resilient, high-performance compute environments for strategic clients across cloud, on-premises, and hybrid deployments. We design, build, and operate large-scale GPU and accelerated-computing infrastructure that supports demanding AI training and inference, simulation, and high-performance compute workloads. Our work spans strategy, architecture, modernization, operations, governance, and continuous improvement across the infrastructure stack. We build reusable operational tools, automation workflows, and platform capabilities that make repeatable infrastructure tasks safer, faster, and more scalable. We collaborate across the technology ecosystem to harness new capabilities, drive business transformation, and deliver dependable services at scale.

Key Responsibilities:

  • Design and implement accelerated-computing infrastructure solutions aligned to system architecture, deployment roadmaps, performance, scalability, resiliency, and governance requirements.

  • Deploy, configure, and operate GPU-based clusters across bare-metal and containerized environments, using workload schedulers and Kubernetes orchestration to support AI training, inference, and high-performance compute workloads.

  • Integrate infrastructure platforms with enterprise systems, data platforms, security frameworks, service-management processes, and governance controls.

  • Design, build, and maintain reusable tools, scripts, self-service capabilities, and automation workflows for infrastructure operations, including provisioning, configuration management, validation, capacity planning, monitoring, incident management, reporting, and recurring remediation.

  • Establish repeatable operational processes for cluster provisioning, configuration management, patching, capacity planning, monitoring, incident response, and lifecycle management.

  • Perform and automate GPU, compute, storage, and network benchmarking and validation; diagnose performance issues across multi-node AI training, inference, and distributed compute workloads.

  • Develop and maintain architecture diagrams, configuration baselines, operational runbooks, and support documentation.

  • Provide technical guidance, troubleshooting, and optimization for GPU clusters supporting AI training, inference, high-performance computing, and multi-node simulation workloads, with emphasis on availability, resiliency, scalability, energy efficiency, and cost management.

Travel may be required for this role. The amount of travel will vary from 25% to 60% depending on business need and client requirements.

Required Skills and Qualifications:

  • Minimum of 5+ years of experience designing, deploying, and managing accelerated-computing infrastructure across on-premises, cloud, and hybrid environments for hyperscaler, neocloud, large enterprise, telecommunications, financial services, manufacturing, and/or retail clients.
  • Minimum of 5+ years of hands-on experience with accelerated-computing platforms, including GPUs, DPUs, and CPUs, high-bandwidth network fabrics, and AI based storage architectures such as parallel file systems, NVMe-oF, etc.
  • Minimum of 5+ years of experience with cluster management, workload scheduling, orchestration, observability, and infrastructure automation, including building operational tools and automation workflows with platforms such as Kubernetes, Slurm, Run:ai
  • Minimum 6 months hands-on experience with Claude Code, AI automation tools, Terraform, Ansible, Python, and Bash scripting.
  • Bachelor's degree or equivalent (minimum 12 years) work experience. (If Associate's Degree, must have minimum 6 years work experience)

Preferred Skills and Qualifications:

  • Experience building AI infrastructure automation and operations tools, AgenticOps practices that enable secure, automated, governed, and reproducible platform operations.
  • Experience developing reusable infrastructure code leveraging Python, platform services, and automation workflows using REST APIs, OpenAPI, JSON/YAML schemas, webhooks, and event-driven integrations.
  • Experience operating large-scale GPU clusters, including capacity management, reliability engineering, change management, and performance validation for AI training, inference, HPC, and enterprise compute workloads.
  • Experience using NVIDIA platform tools and libraries including Base Command Manager (BCM), NGC, NCCL, CUDA-X, NVAIE, Dynamo, benchmarking tools to deploy, tune, profile, and validate cluster performance for training and inference workloads.
  • Experience managing deployments of 1,000+ GPU clusters with infrastructure services enabled for AI training, inference, high-performance, and enterprise compute environments.
  • Design and build experience in running LLMs across AI Cloud platforms from CoreWeave, Nebius, and other specialty providers
  • Knowledge of model deployments, tuning, and troubleshooting performance on AI Infrastructure
  • Industry certifications in accelerated-computing infrastructure, public cloud providers, infrastructure automation, networking, or security are a plus.


Compensation at Accenture varies depending on a wide array of factors, which may include but are not limited to the specific office location, role, skill set, and level of experience. As required by local law, Accenture provides a reasonable range of compensation for roles that may be hired as set forth below.
We anticipate this job posting will be posted until 10/13/2026.
Accenture offers a market competitive suite of benefits including medical, dental, vision, life, and long-term disability coverage, a 401(k) plan, bonus opportunities, paid holidays, and paid time off. See more information on our benefits here:

U.S. Employee Benefits | Accenture


Role Location Annual Salary Range
California $94,400 to $266,300
Cleveland $87,400 to $213,000
Colorado $94,400 to $230,000
District of Columbia $100,500 to $245,000
Illinois $87,400 to $230,000
Maine $80,400 to $196,000
Maryland $94,400 to $230,000
Massachusetts $94,400 to $245,000
Minnesota $94,400 to $230,000
New York $87,400 to $266,300
New Jersey $100,500 to $266,300
Virginia $87,400 to $245,000
Washington $100,500 to $245,000

About Accenture

Accenture is a leading global professional services company that helps the world's leading businesses, governments and other organizations build their digital core, optimize their operations, accelerate revenue growth and enhance citizen services-creating tangible value at speed and scale. We are a talent- and innovation-led company with approximately 791,000 people serving clients in more than 120 countries. Technology is at the core of change today, and we are one of the world's leaders in helping drive that change, with strong ecosystem relationships. We combine our strength in technology and leadership in cloud, data and AI with unmatched industry experience, functional expertise and global delivery capability. Our broad range of services, solutions and assets across Strategy & Consulting, Technology, Operations, Industry X and Song, together with our culture of shared success and commitment to creating 360 value, enable us to help our clients reinvent and build trusted, lasting relationships. We measure our success by the 360 value we create for our clients, each other, our shareholders, partners and communities.

Visit us atwww.accenture.com

What We Believe

We have an unwavering commitment to diversity with the aim that every one of our people has a full sense of belonging within our organization. As a business imperative, every person at Accenture has the responsibility to create and sustain an inclusive environment.

Inclusion and diversity are fundamental to our culture and core values. Our rich diversity makes us more innovative and more creative, which helps us better serve our clients and our communities.Read more here

Requesting An Accommodation

Accenture is committed to providing equal employment opportunities for persons with disabilities or religious observances, including reasonable accommodation when needed. If you are hired by Accenture and require accommodation to perform the essential functions of your role, you will be asked to participate in our reasonable accommodation process. Accommodations made to facilitate the recruiting process are not a guarantee of future or continued accommodations once hired.

If you would like to be considered for employment opportunities with Accenture and have accommodation needs such as for a disability or religious observance, please call us toll free at 1 (877) 889-9009 or send us anemailor speak with your recruiter.

Equal Employment Opportunity Statement

We believe that no one should be discriminated against because of their differences.All employment decisions shall be made without regard to age, race, creed, color, religion, sex, national origin, ancestry, disability status, military veteran status, sexual orientation, gender identity or expression, genetic information, marital status, citizenship status or any other basis as protected by applicable law.Our rich diversity makes us more innovative, more competitive, and more creative, which helps us better serve our clients and our communities.

For details, view a copy of the Accenture Equal Opportunity Statement

Accenture is an EEO and Affirmative Action Employer of Veterans/Individuals with Disabilities.

Accenture is committed to providing veteran employment opportunities to our service men and women.

Other Employment Statements

Applicants for employment in the US must have work authorization that does not now or in the future require sponsorship of a visa for employment authorization in the United States.

Candidates who are currently employed by a client of Accenture or an affiliated Accenture business may not be eligible for consideration.

Job candidates will not be obligated to disclose sealed or expunged records of conviction or arrest as part of the hiring process. Further, at Accenture a criminal conviction history is not an absolute bar to employment.

The Company will not discharge or in any other manner discriminate against employees or applicants because they have inquired about, discussed, or disclosed their own pay or the pay of another employee or applicant. Additionally, employees who have access to the compensation information of other employees or applicants as a part of their essential job functions cannot disclose the pay of other employees or applicants to individuals who do not otherwise have access to compensation information, unless the disclosure is (a) in response to a formal complaint or charge, (b) in furtherance of an investigation, proceeding, hearing, or action, including an investigation conducted by the employer, or (c) consistent with the Company's legal duty to furnish information.

California requires additional notifications for applicants and employees. If you are a California resident, live in or plan to work from Los Angeles County upon being hired for this position, pleaseclick herefor additional important information.

Please read Accenture'sRecruiting and Hiring Statementfor more information on how we process your data during the Recruiting and Hiring process.


What Accenture Federal Services employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom