2

Full Time Failure Analysis Manager Jobs (NOW HIRING)

Requisition ID: 77831 Description As a Failure Analysis Engineering Technician, you are responsible ... Maintain laboratory 5S program, including organization and inventory management of test hardware ...

... management. We're committed to making clean, green energy the primary power source for homes ... We are looking for a Failure Analysis Engineer who loves digging into tough hardware problems ...

next page

Showing results 1-20

Full Time Failure Analysis Manager information

See salary details

$19K

$56K

$98K

How much do full time failure analysis manager jobs pay per year?

As of Jul 23, 2026, the average yearly pay for full time failure analysis manager in the United States is $55,952.00, according to ZipRecruiter salary data. Most workers in this role earn between $38,000.00 and $60,000.00 per year, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive as a Full Time Failure Analysis Manager, and why are they important?

To thrive as a Full Time Failure Analysis Manager, you need a strong background in materials science, engineering, or a related field, often supported by a bachelor’s or master’s degree and relevant industry experience. Familiarity with failure analysis techniques, laboratory equipment (such as SEM, X-ray, and EDX), and reporting tools is essential, along with certifications like Six Sigma or reliability engineering. Strong leadership, analytical thinking, and effective communication are crucial soft skills to manage teams and liaise with cross-functional departments. These skills ensure accurate root cause identification, drive continuous product improvement, and support organizational quality objectives.

What is the difference between Full Time Failure Analysis Manager vs Failure Analysis Engineer?

AspectFull Time Failure Analysis ManagerFailure Analysis Engineer
ResponsibilitiesOversees failure analysis teams, manages projects, and develops strategies for failure preventionPerforms detailed failure investigations, tests, and data analysis on specific components or products
CredentialsBachelor's or Master's in Engineering, certifications like ASQ CQE or CQA often preferredBachelor's or Master's in Engineering or related field, relevant certifications beneficial
Work EnvironmentManagement setting, coordinating teams in labs or manufacturing facilitiesHands-on technical work in labs or testing environments
Industry UsageCommon in manufacturing, electronics, automotive sectorsCommon in electronics, aerospace, automotive industries

The Full Time Failure Analysis Manager focuses on leading teams and strategic planning, while the Failure Analysis Engineer conducts technical investigations and testing. Both roles require engineering credentials and are vital in quality and reliability efforts within manufacturing industries.

What are some common challenges faced by a Full Time Failure Analysis Manager, and how can they be addressed?

A Full Time Failure Analysis Manager often faces challenges such as managing tight deadlines for root cause analysis, balancing multiple high-priority investigations, and ensuring clear communication between engineering, manufacturing, and quality teams. Staying organized and prioritizing tasks based on business impact can help manage workload efficiently. Building strong relationships with cross-functional teams is crucial to gather information quickly and implement effective solutions. Additionally, staying updated with the latest analytical tools and failure analysis techniques can enhance accuracy and speed in resolving issues.

What does a Full Time Failure Analysis Manager do?

A Full Time Failure Analysis Manager oversees the process of investigating product or component failures to identify root causes and recommend corrective actions. They lead teams of engineers and technicians, coordinate testing procedures, analyze data from failed products, and communicate findings to improve quality and prevent future issues. This role is crucial in industries such as electronics, manufacturing, and automotive, where reliability and product performance are critical. The manager also develops protocols, ensures compliance with standards, and collaborates with design and production teams to implement solutions.
More about Full Time Failure Analysis Manager jobs
What are the most commonly searched types of Failure Analysis Manager jobs? The most popular types of Failure Analysis Manager jobs are:
What states have the most Full Time Failure Analysis Manager jobs? States with the most job openings for Full Time Failure Analysis Manager jobs include:
Infographic showing various Full Time Failure Analysis Manager job openings in the United States as of July 2026, with employment types broken down into 85% Full Time, 13% Part Time, and 2% Contract. Highlights an 80% Physical, 5% Hybrid, and 15% Remote job distribution, with an average salary of $55,952 per year, or $26.9 per hour.
Failure Analysis Engineer

Failure Analysis Engineer

Advanced Micro Devices, Inc

Secaucus, NJ • On-site

$84K/yr

Full-time

Posted 16 days ago


Advanced Micro Devices rating

8.6

Company rating: 8.6 out of 10

Based on 13 frontline employees who took The Breakroom Quiz

19th of 144 rated electronics manufacturers


Job description

WHAT YOU DO AT AMD CHANGES EVERYTHING
At AMD, our mission is to build great products that accelerate next-generation computing experiences-from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you'll discover the real differentiator is our culture. We push the limits of innovation to solve the world's most important challenges-striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career.
THE ROLE
Join a highly visible engineering team responsible for bringing up, debugging, and improving next-generation server and rack-scale platforms deployed in data center environments. As a Senior Failure Analysis Engineer, you will be among the first engineers to investigate, troubleshoot, and resolve complex system-level issues on new hardware platforms, partnering closely with design, firmware, validation, manufacturing, quality, and infrastructure teams. This role offers the opportunity to work on cutting-edge technologies while driving root cause analysis, platform reliability, and manufacturing readiness across the product lifecycle. You will play a key role in identifying and resolving complex hardware and firmware interactions, supporting platform bring-up activities, improving debug processes, and helping scale organizational knowledge through technical documentation and AI-assisted engineering workflows. This is an ideal opportunity for an engineer who enjoys solving difficult technical problems, working across multiple disciplines, and making a measurable impact on the success of next-generation data center products.
THE PERSON
The ideal candidate is naturally curious, analytical, and motivated by solving complex technical challenges. They thrive in environments where not every answer is immediately available and are comfortable investigating issues across multiple engineering domains to identify root causes and drive resolution. They demonstrate strong communication skills, effectively manage stakeholder expectations during ongoing investigations, and can clearly communicate technical findings to both engineering and cross-functional teams.
Successful candidates are collaborative, adaptable, and persistent problem-solvers who enjoy learning new technologies, working in highly dynamic environments, and continuously improving how engineering teams operate. They embrace new tools and technologies, including AI-assisted workflows, to improve efficiency, accelerate troubleshooting, and scale technical knowledge across the organization.
KEY RESPONSIBILITIES
  • Perform system-level and rack-level failure analysis across server and data center platforms, driving issues from initial symptom identification through root cause determination and resolution.
  • Debug complex hardware, firmware, and platform-related issues involving CPUs, GPUs, memory, PCIe, networking, power delivery, thermal systems, and firmware interactions.
  • Analyze system logs, platform telemetry, BIOS, BMC, and other diagnostic data to identify failure patterns, isolate root causes, and validate corrective actions.
  • Reproduce failures in lab environments and utilize appropriate diagnostic tools and methodologies to validate fixes and improve overall platform reliability.
  • Partner closely with design, firmware, validation, manufacturing, quality, and infrastructure teams to investigate issues, drive corrective actions, and improve system performance.
  • Support manufacturing and ODM partners by providing technical guidance, failure triage, structured debug processes, and escalation support for complex platform issues.
  • Develop and maintain technical documentation, troubleshooting guides, SOPs, debug methodologies, and knowledge-sharing resources to improve organizational effectiveness.
  • Drive root cause analysis activities and contribute to continuous improvements in reliability, serviceability, manufacturability, and operational readiness.
  • Leverage AI-assisted tools, automation, and knowledge systems to improve troubleshooting efficiency, accelerate investigations, and scale engineering best practices.
  • Serve as a technical resource for cross-functional teams by communicating findings, managing stakeholder expectations, and providing clear recommendations based on data-driven analysis.
  • Contribute expertise in areas such as networking, signal integrity, system architecture, or platform reliability to support the successful deployment of next-generation data center technologies.

PREFERRED EXPERIENCE
  • Experience performing system-level, server-level, or rack-level troubleshooting and failure analysis.
  • Background in hardware debugging, root cause analysis, and complex issue resolution.
  • Familiarity with Linux-based environments, platform logs, and diagnostic workflows.
  • Experience supporting server, data center, networking, storage, or enterprise hardware platforms.
  • Exposure to BIOS, BMC, IPMI, firmware interactions, or platform management technologies.
  • Understanding of networking technologies, signal integrity concepts, power delivery, or high-speed interfaces.
  • Experience collaborating with manufacturing, ODM, validation, quality, or design teams.
  • Ability to develop technical documentation, SOPs, troubleshooting guides, and knowledge-sharing resources.
  • Familiarity with lab debugging equipment and system diagnostic tools.
  • Experience leveraging AI tools, AI agents, automation, or knowledge systems to improve debugging efficiency and technical problem-solving.
  • Willingness to occasionally travel in support of manufacturing, validation, or deployment activities.

ACADEMIC CREDENTIALS
  • Bachelor's or master's degree preferred in Electrical Engineering, Computer Engineering, Systems Engineering, Computer Science, or a related technical field.

LOCATION: Secaucus, NJ (Onsite)
THIS ROLE IS NOT ELEGIBLE FOR VISA SUPPORT
#LI-CS1
Benefits offered are described: AMD benefits at a glance.
AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants' needs under the respective laws throughout all stages of the recruitment and selection process.
AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD's "Responsible AI Policy" is available here.
This posting is for an existing vacancy.

What Advanced Micro Devices employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom