Your credibility with DGX Cloud platform, SRE, networking, and broader NVIDIA security partners is ... senior engineers. You can review a design honestly and recognize when an estimate doesn't add up.
Your credibility with DGX Cloud platform, SRE, networking, and broader NVIDIA security partners is ... senior engineers. You can review a design honestly and recognize when an estimate doesn't add up.
$122K - $161K/yr
Knowledge of SRE principles (observability, SLOs, logging, etc.) Ways to stand out from the crowd: * Experience in a Hyperscale Cloud Service Provider (public-facing or not) * Understanding of ...
$122K - $161K/yr
Knowledge of SRE principles (observability, SLOs, logging, etc.) Ways to stand out from the crowd: * Experience in a Hyperscale Cloud Service Provider (public-facing or not) * Understanding of ...
$94K - $130K/yr
Senior Hardware Engineer Curtiss-Wright is seeking an experienced Senior Hardware Engineer to join ... Apply engineering best practices to improve product performance, reliability, and time-to-market ...
$94K - $130K/yr
Senior Hardware Engineer Curtiss-Wright is seeking an experienced Senior Hardware Engineer to join ... Apply engineering best practices to improve product performance, reliability, and time-to-market ...
$94K - $130K/yr
Senior Hardware Engineer - Remote Curtiss-Wright is seeking an experienced Senior Hardware Engineer ... Apply engineering best practices to improve product performance, reliability, and time-to-market ...
$94K - $130K/yr
Senior Hardware Engineer - Remote Curtiss-Wright is seeking an experienced Senior Hardware Engineer ... Apply engineering best practices to improve product performance, reliability, and time-to-market ...
$122K - $161K/yr
We are looking for a Senior Software Engineer to help build NeMo Platform, NVIDIA's product for ... Improving reliability, observability, debuggability, and performance across NeMo Platform, SDKs ...
$122K - $161K/yr
We are looking for a Senior Software Engineer to help build NeMo Platform, NVIDIA's product for ... Improving reliability, observability, debuggability, and performance across NeMo Platform, SDKs ...
$89K - $116K/yr
... reliability standards, standard industry practices, etc. * Provide design review and advice on any ... Senior Engineer : * 7+ years of a combination of technical design experience. This can include ...
$89K - $116K/yr
... reliability standards, standard industry practices, etc. * Provide design review and advice on any ... Senior Engineer : * 7+ years of a combination of technical design experience. This can include ...
$104K - $143K/yr
As a Senior Systems Engineer your duties include, but are not limited to the following: * Conduct ... cost, reliability, maintainability, risk, and schedule. * Offer technical recommendations through ...
$104K - $143K/yr
As a Senior Systems Engineer your duties include, but are not limited to the following: * Conduct ... cost, reliability, maintainability, risk, and schedule. * Offer technical recommendations through ...
$104K - $143K/yr
As a Senior Systems Engineer your duties include, but are not limited to the following: * Provide ... reliability/maintainability actions, and worldwide user support * Assist in designing, planning ...
$104K - $143K/yr
As a Senior Systems Engineer your duties include, but are not limited to the following: * Provide ... reliability/maintainability actions, and worldwide user support * Assist in designing, planning ...
$104K - $142K/yr
Cloud Foundations Reliability (CFR) is part of NVIDIA's Global Network Infrastructure (GNI ... We are looking for a hands-on senior engineer to own the lifecycle and automation of the Kubernetes ...
$104K - $142K/yr
Cloud Foundations Reliability (CFR) is part of NVIDIA's Global Network Infrastructure (GNI ... We are looking for a hands-on senior engineer to own the lifecycle and automation of the Kubernetes ...
$104K - $143K/yr
We are seeking a Senior HPC & Quantum Systems Engineer to help architect, deploy, and operate a ... Be responsible for the reliability, performance, administrations, and lifecycle management of the ...
$104K - $143K/yr
We are seeking a Senior HPC & Quantum Systems Engineer to help architect, deploy, and operate a ... Be responsible for the reliability, performance, administrations, and lifecycle management of the ...
$122K - $161K/yr
Create monitoring and alerting solutions for system health and reliability. Cloud & DevOps * Deploy and manage applications within cloud environments such as: Azure, AWS, Google Cloud, etc. * Support ...
$122K - $161K/yr
Create monitoring and alerting solutions for system health and reliability. Cloud & DevOps * Deploy and manage applications within cloud environments such as: Azure, AWS, Google Cloud, etc. * Support ...
... to the C2ISR/HBG Senior Materiel Leader. EPASS support is required to execute the major GNS ... Support engineering activities related to logistics, sustainment, reliability, maintainability, and ...
... to the C2ISR/HBG Senior Materiel Leader. EPASS support is required to execute the major GNS ... Support engineering activities related to logistics, sustainment, reliability, maintainability, and ...
Drive tradeoffs across bandwidth, latency, power, area, timing, scalability, reliability, security ... What We Need To See: * BS, MS, or PhD in Electrical Engineering, Computer Engineering, Computer ...
Drive tradeoffs across bandwidth, latency, power, area, timing, scalability, reliability, security ... What We Need To See: * BS, MS, or PhD in Electrical Engineering, Computer Engineering, Computer ...
$200K - $275K/yr
Lead and oversee quality management systems to maintain product reliability and customer trust ... Partner closely with R&D, engineering, supply chain, quality, and regional operations teams to ...
$200K - $275K/yr
Lead and oversee quality management systems to maintain product reliability and customer trust ... Partner closely with R&D, engineering, supply chain, quality, and regional operations teams to ...
Establish reliability, security, validation, and left-shift strategies that reduce risk before hardware reaches production environments. * Mentor senior engineers and technical leads, raising the ...
Establish reliability, security, validation, and left-shift strategies that reduce risk before hardware reaches production environments. * Mentor senior engineers and technical leads, raising the ...
$108K - $147K/yr
NVIDIA is seeking an experienced software engineer to join the Cloud Foundations Automation team ... Build observability and security capabilities that improve the reliability and resilience of our ...
$108K - $147K/yr
NVIDIA is seeking an experienced software engineer to join the Cloud Foundations Automation team ... Build observability and security capabilities that improve the reliability and resilience of our ...
$67.25 - $90/hr
Collaborate with NVIDIA product, engineering and partners teams to understand NVIDIA's reference ... Review the operational efficiency, reliability, and readiness of data center and hardware elements ...
New
$67.25 - $90/hr
Collaborate with NVIDIA product, engineering and partners teams to understand NVIDIA's reference ... Review the operational efficiency, reliability, and readiness of data center and hardware elements ...
New
$112K - $257K/yr
... senior role on delivered systems * Experience in a server-side language, such as Java, .NET, C# ... Ability to translate mission and reliability requirements into scalable, resilient service designs
$112K - $257K/yr
... senior role on delivered systems * Experience in a server-side language, such as Java, .NET, C# ... Ability to translate mission and reliability requirements into scalable, resilient service designs
Raytheon brings the strength of more than 100 years of experience and renowned engineering ... This team of senior architects leverages common frameworks, processes, and technology patterns to ...
Raytheon brings the strength of more than 100 years of experience and renowned engineering ... This team of senior architects leverages common frameworks, processes, and technology patterns to ...
$63.75 - $82/hr
Data Architect (S/4HANA) Sr. Manager, Enterprise Architecture & Strategy - Individual Contributor ... Experience in data architecture/engineering with endtoend ownership of enterprise data platforms.
$63.75 - $82/hr
Data Architect (S/4HANA) Sr. Manager, Enterprise Architecture & Strategy - Individual Contributor ... Experience in data architecture/engineering with endtoend ownership of enterprise data platforms.
Senior Reliability Engineer information
What does a senior reliability engineer do?
What are the key skills and qualifications needed to thrive as a senior reliability engineer?
What are some common challenges faced by senior reliability engineers, and how are they typically addressed within the team?
What is the difference between Senior Reliability Engineer vs Reliability Engineer?
| Aspect | Senior Reliability Engineer | Reliability Engineer |
|---|---|---|
| Credentials | Typically requires 5+ years experience, certifications like CRE or Six Sigma | Entry to mid-level, often with 2-4 years experience, similar certifications |
| Work Environment | Designs and oversees reliability programs, leads projects | Performs analysis, supports reliability improvements |
| Industry Usage | Used across manufacturing, energy, aerospace | Common in same industries, often as a stepping stone to senior roles |
The main difference between a Senior Reliability Engineer and a Reliability Engineer lies in experience, leadership responsibilities, and scope of work. Senior Reliability Engineers typically lead projects and develop strategies, while Reliability Engineers focus on analysis and supporting reliability initiatives. Both roles are vital in ensuring equipment and system dependability across industries.
How much do senior reliability engineers get paid?
What are popular job titles related to Senior Reliability Engineer jobs in FM?
For Senior Reliability Engineer jobs in FM, the most frequently searched job titles are:

Senior Engineering Manager, Infrastructure Security Engineering - DGX Cloud
On-site, Remote
Full-time
Posted 9 days ago
Key responsibilities
Manage, coach, and develop a distributed team of senior security and infrastructure engineers.
Partner with engineers to set the roadmap for security control plane, foundational security services, and autonomous security operations.
Run a planning process focused on delivering risk elimination outcomes and assemble teams around specific work efforts.
Nvidia rating
9.6
Based on 18 frontline employees who took The Breakroom Quiz
Job description
NVIDIA DGX Cloud is the AI supercomputing-as-a-service substrate designed to power the next generation of AI and industrial-scale breakthroughs. As an Engineering Manager within our Infrastructure Security Engineering organization, you will enable the engineers who secure a GPU fleet numbering in the hundreds of thousands. Your team does not write reports about risk. It engineers whole classes of risk out of existence, and delivers the result as paved roads the rest of DGX Cloud will happily take.
What You Will Be Doing:
Grow the Team: Manage, coach, and develop a distributed team of senior security and infrastructure engineers. You own hiring, growth, and performance, and you build a bench deep enough that our capability never rests on one person. Heroics signal a system that needs fixing, so you protect sustainable pace.
Set Technical Direction: Partner with your engineers on the roadmap for our security control plane, our foundational security services, and the path toward autonomous security operations. Leadership owns the mission and the strategy; the team owns the execution.
Deliver Outcomes, Not Tickets: Run a lightweight planning rhythm where the team commits to a result and the reason it matters, then reflects openly on what to change next. We are measured by the risks we eliminate, not the tickets we close.
Assemble Teams Around the Work: Staff each effort with the smallest group who can deliver it, and a clear owner accountable for landing it. Teams form around the work and dissolve when it's done. Make that motion fast, fair, and a real growth opportunity.
Raise the Engineering Bar: Hold a high standard for what "done" means: tested code, observability, operational readiness, and documentation someone who didn't build the system can operate from. AI-assisted work ships with real human ownership behind it.
Communicate Risk Upward: Translate a complex, fast-moving risk picture into something executives can act on. Declare the unknowns honestly, turn risk assessment into measurable themes, and make trade-offs legible to the people funding them.
Multi-Functional Collaboration: Build the partnerships that make the work land. We operate horizontally across vertically organized teams, with influence rather than authority. Your credibility with DGX Cloud platform, SRE, networking, and broader NVIDIA security partners is a delivery mechanism, not a nicety.
What We Need to See:
We are looking for high-caliber engineering leaders with deep spikes of expertise in a few of these areas and the intellectual curiosity to dive into the rest. If your experience aligns with the core of this role-building and growing teams that engineer risk out of existence-and you can show us how, we want to hear from you!
Engineering Management: Experience (7+ years) managing engineers, within a broader career (14+ years) across software engineering, SRE, infrastructure, or security. You have grown people, not only shipped projects.
Technical Credibility: Enough hands-on depth in infrastructure, distributed systems, or security engineering to earn the trust of senior engineers. You can review a design honestly and recognize when an estimate doesn't add up. You don't need to be the best engineer in the room; you need to build the room.
Hiring and Developing Talent: A history of attracting, hiring, and growing excellent engineers, including people far more expert than you in their domain. You build interview and calibration practices that make hiring repeatable and fair.
Delivery Under Ambiguity: You can take a high-stakes, loosely defined mandate and turn it into sequenced, measurable delivery. You are comfortable proving a model on a few systems first, then using that success to expand.
Cloud-Native and Security Fluency: Working understanding of cloud-native architecture, container orchestration (Kubernetes), identity and access, policy enforcement, and vulnerability management-enough to build it and ask the questions that matter.
Influence Without Authority: Experience delivering through partner teams you do not own, in environments where relationships, not org charts, determine whether the work ships.
Foundation: Bachelor's degree in Computer Science, Engineering, or a related technical field (or equivalent experience).
Ways To Stand Out from the Crowd:
HPC/AI Infrastructure: Experience running teams that secure or operate high-performance computing environments, large GPU fleets, or multi-tenant platforms at cloud scale.
Security as an Internal Product: You have delivered security or infrastructure capabilities with real adoption metrics, where teams took the paved road because it was the easiest path.
Building a Function from Scratch: Experience standing up a team, a practice, or an engineering culture where none existed-charter, agreements, rituals, and all.
Scaling Without Dilution: A history of growing a team quickly while the bar went up rather than down. How have you scaled a team and raised the standard at the same time?
You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.About Nvidia
Sourced by ZipRecruiter
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.
Industry
Computer and electronic product manufacturing
Company size
10,000+ Employees
Headquarters location
Santa Clara, CA, US