Failure Analysis Engineer - Manufacturing
SuperMicroComputer · San Jose, California, United States · On-site
Pay: USD 90,000 – 125,000 a year
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
Job Req ID: 30223
About Supermicro:
Supermicro® is a Top Tier provider of advanced server, storage, and networking solutions for Data Center, Cloud Computing, Enterprise IT, Hadoop/ Big Data, Hyperscale, HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon Valley Top 50 technology firms. Our unprecedented global expansion has provided us with the opportunity to offer a large number of new positions to the technology community. We seek talented, passionate, and committed engineers, technologists, and business leaders to join us.
Job Summary:
We are seeking a highly skilled and experienced Failure Analysis Engineer - Manufacturin to join our team at Supermicro. In this pivotal role, you will lead critical investigations into product failures, utilizing your deep expertise in failure analysis methodologies to drive root cause identification and implement corrective actions that enhance product reliability and performance.
Essential Duties and Responsibilities:
Investigate hardware, software, and network-related failures in servers and related systems to identify root causes
Run tests to reproduce failure modes and validate fixes.
Perform reporting on failure analysis and action items.
Work with design, manufacturing, engineer, and support teams to implement design changes and process improvements
Perform hardware repairs, component replacements, and system reconfigurations to restore service quickly
Provides technical support to internal and external customers (travel required)
Provide mentoring, coaching and development for junior-level resources
Qualifications:
Bachelor’s or Master’s degree in Engineering (Electrical, Computer, Mechanical, or related), or a related technical field
3+ years experiences in related roles, with hands-on experience in server/rack systems, data centers, or failure analysis labs
Strong understanding of server /rack architecture, operation system and GPU systems
Familiar with oscilloscopes and strong knowledge of analysis flows and methodology.
Ability to troubleshoot complex hardware/software issues and think critically under pressure.
Circuit board design background experience a plus
Physical Requirements and Work Conditions:
The physical demands and work conditions described here are representative of those that must be met by an employee to successfully perform the essential functions of this job:
Able to stand, sit, walk, bend, stoop, reach, lift, and carry items.
Perform tasks that require standing and walking for long periods up to entire shift.
Lift, carry, push and pull in an excess of 25lb and up to 40lbs. Able to utilize a "buddy system" for moving objects more than 50 lbs
(For applicable positions: Must be able to operate pallet jack and forklift efficiently and safely.)
Must be able to wear safety shoes, earmuffs or other required PPE as needed.
Primarily works indoors with controlled climate…