Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
Scoutfield Logo

See all jobs on Scoutfield

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Supermicro

Failure Analysis Engineer – Manufacturing

Supermicro

. Investigate hardware, software, and network-related failures in servers and related systems to identify root causes .

Posted 9/15/2026full-timeSan Jose • California • United StatesMid-LevelSenior💰 $90,000 - $125,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in troubleshooting hardware and software issues within server and rack systems, with a strong focus on failure analysis and process improvements. Capable of mentoring junior resources while providing technical support to both internal and external customers.

Highest-signal resume keywords
Bachelor’s Degree In Engineering3+ Years Of Experience In Related RolesServer/Rack Architecture UnderstandingFailure Analysis MethodologyTechnical Support Experience

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Troubleshooting Hardware IssuesTroubleshooting Software IssuesFailure AnalysisCircuit Board DesignServer/Rack Systems ExperienceData Center OperationsOperating Systems KnowledgeGPU Systems UnderstandingAnalysis FlowsHands-On Technical Skills
Soft Skills
Critical Thinking Under PressureMentoringCoachingCollaborationCommunication
Tools & Technologies
OscilloscopePPE Equipment
Industry Keywords
Hardware RepairsComponent ReplacementsSystem ReconfigurationsDesign ChangesProcess Improvements

About the role

Key responsibilities & impact
  • Investigate hardware, software, and network-related failures in servers and related systems to identify root causes
  • Run tests to reproduce failure modes and validate fixes
  • Report on failure analysis and action items
  • Work with design, manufacturing, engineering, and support teams to implement design changes and process improvements
  • Perform hardware repairs, component replacements, and system reconfigurations to restore service quickly
  • Provide technical support to internal and external customers, including required travel
  • Mentor, coach, and develop junior-level resources

Requirements

What you’ll need
  • Bachelor’s or Master’s degree in Engineering (Electrical, Computer, Mechanical, or related), or a related technical field
  • 3+ years of experience in related roles
  • Hands-on experience in server/rack systems, data centers, or failure analysis labs
  • Strong understanding of server/rack architecture, operating systems, and GPU systems
  • Familiarity with oscilloscopes
  • Strong knowledge of analysis flows and methodology
  • Ability to troubleshoot complex hardware/software issues and think critically under pressure
  • Circuit board design background experience is a plus
  • Ability to stand, sit, walk, bend, stoop, reach, lift, and carry items
  • Ability to perform tasks requiring standing and walking for long periods, up to an entire shift
  • Ability to lift, carry, push, and pull in excess of 25 lb and up to 40 lb
  • Ability to use a buddy system for moving objects over 50 lb
  • Ability to wear safety shoes, earmuffs, or other required PPE as needed
  • Regular in-office attendance during standard working hours
  • Ability to work primarily indoors in controlled climate conditions

Benefits

Comp & perks
  • Comprehensive benefits package
  • Potential eligibility for bonus programs
  • Potential eligibility for equity award programs
  • In-office collaboration and participation in team meetings, training sessions, and other on-site activities