Description
Join NVIDIA's datacenter product engineering team in Operations and drive failure analysis and debug efforts during Mass Production (MP) phase. As a Senior System Debug Engineer, you will collaborate with top minds to ensure flawless manufacturing of datacenter GPU products.
Responsibilities:
- Perform failure analysis on GPU baseboards and servers at various levels.
- Analyze logs and failures spanning Hardware, Software, and Firmware, proposing debug and mitigation strategies.
- Build experiments, collect and analyze data for Failure Analysis root cause.
- Engage in Build for Test, Manufacturing (DFx) enabling efforts and deliver products on schedule.
- Provide root cause and corrective action plans, writing clear reports.
- Develop debug guides for partner teams and customers.
- Communicate effectively with industry vendors, engineers, and sales teams.
Requirements:
- 12+ years of experience in a related field.
- Bachelor’s or Master’s degree in Electrical Engineering or related field.
- Excellent failure analysis or debug experience on motherboards, graphic cards, servers, or datacenter products.
- Demonstrated knowledge in enabling DFx requirements.
- Strong skills in Hardware, Software, Component, Process, Test, or Validation.
- Strong negotiation, organization, and time management skills.
- Familiarity with characterization equipment like oscilloscopes and analyzers.
- Problem-solving mentality and ability to work independently and in a team.
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting:
https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/Senior-Debug-System-Engineer--Datacenter_JR2019809