Job Summary:The Repair Engineer is responsible for AI server repair process support, failure analysis, repair validation, rework guidance, defect reduction, and repair yield improvement for production and sustaining operations. This role supports troubleshooting, repair instruction development, component-level or system-level rework coordination, repair data analysis, and corrective action follow-up. The Repair Engineer works closely with Repair Technicians, Test Engineering, Diagnostic Engineering, Manufacturing, Quality, Materials, Engineering to reduce repair cycle time, improve repair quality, and support production recovery.
Duties and Responsibilities:- Support AI server repair operations, process readiness, troubleshooting, and repair issue resolution across assigned production areas.
- Review failed units, test logs, diagnostic findings, repair history, symptoms, and defect data to guide component-level and system-level repair actions.
- Develop and maintain repair procedures, rework instructions, troubleshooting guides, repair checklists, validation criteria, repair records, and failure-closure documentation.
- Monitor repair yield, cycle time, repeat failures, no-trouble-found results, scrap risk, rework quality, and repair aging; drive containment, corrective actions, and verification of effectiveness.
- Partner with Test Engineering, Diagnostic Engineering, Manufacturing, Quality, Materials, and Engineering to resolve repair blockers, train technicians, and improve repair accuracy, quality, and material recovery.
Education:- Bachelor's degree in Electrical Engineering, Electronics Engineering, Manufacturing Engineering, or a related technical field required.
- Relevant coursework, certification, or hands-on training in electronics repair, server hardware, rework, diagnostics, or manufacturing systems preferred.
Experience:- 4-7 years of experience in repair engineering, failure analysis, electronics repair, manufacturing test, diagnostics, or related technical fields.
- Demonstrated experience with repair procedures, rework instructions, log review, retest validation, component replacement, system-level troubleshooting, and defect containment.
- Experience supporting AI server, server hardware, rack-level systems, Linux, BMC, BIOS, firmware, networking, storage, GPU, power, thermal, or hardware debug environments preferred.