1

On Call Container Graphics Jobs (NOW HIRING)

Senior Site Reliability Engineer - HPC

Durham, NC · On-site

$55 - $73.25/hr

NVIDIA has been transforming computer graphics and accelerated computing for over 25 years. They ... on-call, incident reviews, assist in root cause identification, and produce high-quality RCA ...

HPC Support Engineer

$122K - $162K/yr

This position is expected to participate in an on-call rotation. What You'll Do * Serve as a senior ... Experience with virtualization and container (Docker, Kubernetes) technologies. * Experience with ...

Engineer III

Mason, OH · Hybrid

$54 - $72.75/hr

Collaborates with engineers and graphic designers, analyzes and classifies complex change request ... Provides on call support and monitors the system and identifies system deficiencies. Minimum ...

Engineer III

Mason, OH · Hybrid

$54 - $72.75/hr

Collaborates with engineers and graphic designers, analyzes and classifies complex change request ... Provides on call support and monitors the system and identifies system deficiencies. Minimum ...

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than ... Participate in on-call, incident reviews, assist in root cause identification, and produce high ...

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than ... Participate in on-call, incident reviews, assist in root cause identification, and produce high ...

On Call Container Graphics information

See salary details

$19

$34

$48

How much do on call container graphics jobs pay per hour?

As of Aug 8, 2026, the average hourly pay for on call container graphics in the United States is $34.67, according to ZipRecruiter salary data. Most workers in this role earn between $27.88 and $41.83 per hour, depending on experience, location, and employer.
What are the most commonly searched types of Container Graphics jobs? The most popular types of Container Graphics jobs are:
Infographic showing various On Call Container Graphics job openings in the United States as of August 2026, with employment types broken down into 1% Locum Tenens, 1% As Needed, 77% Full Time, 15% Part Time, and 6% Contract. Highlights an 96% Physical, 1% Hybrid, and 3% Remote job distribution, with an average salary of $72,122 per year, or $34.7 per hour.

Senior Site Reliability Engineer - HPC

NVIDIA

Durham, NC • On-site

$55 - $73.25/hr

Full-time

Re-posted 13 days ago


Nvidia rating

9.6

Company rating: 9.6 out of 10

Based on 17 frontline employees who took The Breakroom Quiz

8th of 242 rated software companies


Job description

Job Summary:
NVIDIA has been transforming computer graphics and accelerated computing for over 25 years. They are seeking a Senior Site Reliability Engineer to join their Compute Farm team, responsible for ensuring the reliability and performance of critical systems while leveraging AI to deliver innovative solutions.
Responsibilities:
• Own SRE solutions end‑to‑end, from design and implementation to operation and continuous improvement, ensuring they integrate cleanly with HPC schedulers, storage, and network fabrics.
• Use IaC(Infrastructure‑as‑Code) and config management to standardize and automate provisioning everywhere.
• Deliver solutions in a globally distributed, multi‑cloud hybrid environment – On‑prem, AWS, GCP, and OCI.
• Design for failure with redundancy, failure domains, progressive delivery, and strict change control.
• Ensure the highest level of uptime and Quality of Service (QoS) for internal customers through operational excellence.
• Conduct capacity management and planning to meet ongoing operational needs.
• Detects performance issues and recommends solutions to maintain world‑class service quality.
• Collaborate with various teams in a fast‑paced environment to ensure seamless project completion.
• Participate in on-call, incident reviews, assist in root cause identification, and produce high-quality RCA reports.
Qualifications:
Required:
• B.S. degree in Computer Science or related technical field (or equivalent experience) with 5+ years professional experience building and supporting critical services.
• Experience supporting large‑scale HPC clusters using Slurm, LSF or Kubernetes clusters, including setup, tuning, and troubleshooting.
• Proficiency in modern CI/CD techniques, and Infrastructure as Code (IaC) for managing services.
• Strong experience crafting large-scale infrastructure platforms for automated host lifecycle management, fleet reliability/auto-healing, E2E observability or data-driven operations (AIOps/ML-driven signals) that materially reduce manual intervention.
• Proficient in monitoring, metrics, container management, and log collection tools.
• 5+ years of coding/scripting experience in at least two high‑level programming languages such as Python, Go, Perl, or Ruby.
• Mentored other engineers and influenced technical direction through design reviews, architecture documents, and strong partnership with product and leadership.
• Creative problem solver with excellent debugging skills and strong communication and documentation abilities.
Preferred:
• Published technical write‑ups or talks (conference presentations, meetups, engineering blogs) that deep‑dive into real‑world reliability, observability, or large‑scale HPC/SRE problems and their solutions.
• Maintainer or co‑maintainer responsibilities for an open source component used in production (plugins, operators, exporters, controllers, or SDKs) at large scale.
Company:
NVIDIA is a computing platform company operating at the intersection of graphics, HPC, and AI. Founded in 1993, the company is headquartered in Santa Clara, USA, with a team of 10001+ employees. The company is currently Late Stage.

What Nvidia employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Nvidia logo

About Nvidia

Sourced by ZipRecruiter

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.

Industry

Computer and electronic product manufacturing

Company size

10,000+ Employees

Headquarters location

Santa Clara, CA, US

Year founded

1993