1

Sr Failure Analysis Engineer Jobs (NOW HIRING)

Failure Analysis Engineer

Maynard, MA · On-site

$77K - $98K/yr

Working closely with failure analysis engineers and cross-functional partners, the team investigates customer-reported issues, identifies root causes, and continuously improves failure analysis ...

Working closely with failure analysis engineers and cross-functional partners, the team investigates customer-reported issues, identifies root causes, and continuously improves failure analysis ...

NPI build defective boards debugging, analysis and repair. * Work with product engineering team to figure out production issue (Including SMT and PCBA assembly). * Come out failure analysis report ...

Failure Analysis Engineer

Fremont, CA · On-site

$142K - $169K/yr

Position Summary We are seeking a highly analytical Failure Analysis Engineer to support the investigation of hardware failures in rack systems, server platforms, and data center infrastructure ...

Failure analysis laboratory process engineer in Fayetteville, AR. Individual will provide support for customer returns, new product introduction and continual improvement for Wolfspeed SiC-based ...

Description Position at Samtec, Inc Samtec is seeking a Failure Analysis Engineer to join the Engineering Support Group (ESG) team based in Colorado Springs, CO. If your favorite part of engineering ...

Failure Analysis Engineer

Fremont, CA · On-site

$142K - $169K/yr

Position Summary We are seeking a highly analytical Failure Analysis Engineer to support the investigation of hardware failures in rack systems, server platforms, and data center infrastructure ...

Samtec, Inc Samtec is seeking a Failure Analysis Engineer to join the Engineering Support Group (ESG) team based in Colorado Springs, CO. If your favorite part of engineering is rolling up your ...

Descripcion Puesto en Samtec, Inc Samtec is seeking a Failure Analysis Engineer to join the Engineering Support Group (ESG) team based in Colorado Springs, CO. If your favorite part of engineering is ...

next page

Showing results 1-20

Sr Failure Analysis Engineer information

See salary details

$59.5K

$126.6K

$183.5K

How much do sr failure analysis engineer jobs pay per year?

As of Jul 24, 2026, the average yearly pay for sr failure analysis engineer in the United States is $126,557.00, according to ZipRecruiter salary data. Most workers in this role earn between $104,500.00 and $143,500.00 per year, depending on experience, location, and employer.

What are the key skills and qualifications needed to thrive as a Sr Failure Analysis Engineer, and why are they important?

To thrive as a Sr Failure Analysis Engineer, you need a strong background in materials science, electronics, and failure analysis methodologies, usually supported by a degree in engineering or a related field. Expertise with analytical tools such as SEM, TEM, FIB, and industry-standard failure analysis software is typically required, along with knowledge of quality standards like Six Sigma or IPC. Critical thinking, problem-solving, and clear communication are key soft skills for effectively diagnosing issues and presenting findings to cross-functional teams. These competencies are vital for accurately identifying root causes, preventing future failures, and driving continuous product improvement.

What is the difference between Sr Failure Analysis Engineer vs Failure Analysis Engineer?

AspectSr Failure Analysis EngineerFailure Analysis Engineer
Required CredentialsBachelor's or Master's in Engineering, certifications like ASQ FAI or relatedBachelor's or Master's in Engineering, similar certifications
Work EnvironmentManufacturing, electronics, aerospace, or automotive industriesSame industries, often in R&D or quality departments
Employer & Industry UsageUsed in large corporations, tech firms, manufacturing plantsCommon in similar settings, often as entry to mid-level roles

The main difference between a Sr Failure Analysis Engineer and a Failure Analysis Engineer lies in experience and responsibility. The senior role typically involves more complex analysis, mentorship, and project leadership, while the standard role focuses on executing failure investigations under supervision. Both roles require similar credentials and work in comparable environments, but the senior position indicates a higher level of expertise and accountability.

What are some common challenges faced by a Sr Failure Analysis Engineer when working with cross-functional teams?

A Sr Failure Analysis Engineer often collaborates with design, manufacturing, and quality teams to identify the root causes of product failures. One common challenge is effectively communicating complex technical findings to non-engineering stakeholders and ensuring recommendations are clearly understood and actionable. Additionally, balancing multiple urgent failure investigations while maintaining detailed documentation and meeting tight deadlines can be demanding. Building strong cross-functional relationships and proactively sharing insights helps facilitate smoother collaboration and more successful problem resolution.

What does a Sr Failure Analysis Engineer do?

A Sr Failure Analysis Engineer investigates and determines the root causes of product or component failures in manufacturing or field environments. They use various analytical tools and techniques, such as microscopy, electrical testing, and material analysis, to examine failed products and recommend corrective actions. Their work helps improve product reliability, prevent recurring issues, and support quality assurance efforts. They also prepare detailed reports and collaborate with design, manufacturing, and quality teams to implement solutions.
More about Sr Failure Analysis Engineer jobs
What cities are hiring for Sr Failure Analysis Engineer jobs? Cities with the most Sr Failure Analysis Engineer job openings:
What states have the most Sr Failure Analysis Engineer jobs? States with the most job openings for Sr Failure Analysis Engineer jobs include:
Infographic showing various Sr Failure Analysis Engineer job openings in the United States as of July 2026, with employment types broken down into 79% Full Time, 12% Part Time, and 9% Contract. Highlights an 76% Physical, 4% Hybrid, and 20% Remote job distribution, with an average salary of $126,557 per year, or $60.8 per hour.
Senior Failure Analysis Engineer - Test Development

Senior Failure Analysis Engineer - Test Development

Advanced Micro Devices, Inc

Secaucus, NJ

Full-time

Posted 18 days ago


Advanced Micro Devices rating

8.6

Company rating: 8.6 out of 10

Based on 13 frontline employees who took The Breakroom Quiz

19th of 144 rated electronics manufacturers


Job description


WHAT YOU DO AT AMD CHANGES EVERYTHING 

At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond.  Together, we advance your career.  



The ROLE:

The Quality Engineering team is looking for an experienced Senior Failure Analysis Engineer - Test Development  to create advanced test methods that surface elusive failures in GPU accelerator platforms. This role is centered on designing custom execution flows that go beyond standard validation, using stress-based scenarios, VPOD environments, AI/ML workloads, and adaptive test logic to make hard-to-capture issues observable and actionable. The engineer will expand failure analysis capability across lab, factory, and customer-return cases by building test content that improves repeatability, shortens debug cycles, and increases confidence in root cause findings. They will also help shape intelligent test systems that use internal engineering knowledge and live model inference to guide execution decisions in real time. Working across FA, validation, firmware, diagnostics, and data teams, this person will help convert unclear symptoms into testable conditions that accelerate resolution. 

THE PERSON: 

The ideal candidate is inventive, methodical, and technically versatile, with a strong instinct for designing experiments that reveal behavior hidden under normal test conditions. They are comfortable navigating hardware, firmware, software, and system-level interactions, and know how to choose the right levers—environment, timing, workload composition, instrumentation, or automation—to provoke meaningful behavior. They are effective in VPOD-based test environments, capable of using model-driven compute activity as part of system stimulation, and confident building AI-enabled workflows that draw from team-specific knowledge during execution. Just as importantly, they can turn messy observations into disciplined experiments, communicate clearly across teams, and document approaches in a way others can reuse. 

KEY RESPONSIBILITIES: 

  • Architect targeted test methods for hard-to-capture platform behaviors across GPU, server, and rack-scale environments. 

  • Invent new workload patterns, sequencing approaches, and stress combinations that reveal conditions not covered by conventional diagnostics. 

  • Build and maintain VPOD-based environments that support scalable experimentation, long-duration execution, and controlled reproduction studies. 

  • Use inference and training activity as system stimuli to probe platform limits, timing sensitivities, and failure-prone operating regions. 

  • Develop automation, scripting, and orchestration tools to launch workloads, monitor execution, collect logs, and analyze results at scale across Windows and Linux environments. 

  • Interpret telemetry, logs, and observed signatures to refine experiments, isolate trigger conditions, and improve confidence in reproduced behavior. 

  • Create AI-enabled execution flows that use internal FA knowledge and live inference to guide test branching, detect emerging patterns, and support faster triage decisions. 

  • Partner closely with FA, validation, diagnostics, firmware, and manufacturing teams to translate vague symptoms or sporadic field issues into targeted and repeatable test content. 

  • Document workload intent, test methods, reproduction conditions, and findings clearly so they can be reused across teams and incorporated into future FA workflows. 

  • Drive continuous improvement of test development methods, workload libraries, and failure reproduction strategies to expand FA coverage and reduce time to root cause. 

PREFERRED EXPERIENCE: 

  • Proven track record of developing custom test methodologies for intermittent, low-occurrence, or otherwise difficult-to-observe failure modes. 

  • Strong foundation in GPU and server platform behavior, including system stress interactions, concurrency effects, and stability characterization. 

  • Demonstrated ability to build, run, and optimize VPOD environments and related infrastructure for large-scale FA or validation test execution. 

  • Hands-on familiarity with inference and training environments, including their use as controllable system stressors in platform investigation. 

  • Proficient in Python, shell scripting, and automation development for workload launch, orchestration, telemetry capture, and post-run analysis. 

  • Ability to interpret system data and debug artifacts to uncover meaningful signals and guide the next experimental step. 

  • Familiarity with diagnostics, firmware interactions, drivers, and hardware/software boundaries that influence failure behavior under stress workloads. 

  • Experience building AI-enabled test systems that incorporate internal engineering knowledge and support real-time inference during execution. 

  • Strong communication, documentation, collaboration, and presentation skills, with the ability to explain complex reproduction strategies and findings across technical teams. 

  • Experience with GPU data center infrastructure, AI/ML technologies, and non-standard workload development is a strong plus. 

ACADEMIC CREDENTIALS: 

  • Bachelor’s degree in Electrical Engineering, Computer Engineering, Computer Science, or a related field. 

 

This role is not eligible for Visa sponsorship 

#LI-LB1



Benefits offered are described:  AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law.   We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position.  AMD’s “Responsible AI Policy” is available here.

 

This posting is for an existing vacancy.

Qualifications:

Benefits offered are described:  AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law.   We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position.  AMD’s “Responsible AI Policy” is available here.

 

This posting is for an existing vacancy.

Education:UNAVAILABLEEmployment Type: FULL_TIME

What Advanced Micro Devices employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom