Sr. Sys/Test Validation Engineer
AMD
- Location
- Hsinchu, Taiwan
- Employment
- Full Time
- Work model
- On-Site
- Level
- Senior
- Posted
- 1h ago
Skills
About this role
ADVANCE YOUR CAREER. ADVANCE THE WORLD. At AMD, we believe technology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future. Whether you’re designing next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger — technology that moves the world forward. Join us and, together, we’ll advance your career.
THE ROLE
The Board and System Test Engineering group supports the full product life cycle , with this position focused on tester equipment and server-based test infrastructure. Activities include defining design-for-test and infrastructure requirements, developing and qualifying tester configurations, supporting new product introduction (NPI), and sustaining reliable, scalable test capacity through mass production. Responsibilities include ownership of factory tester equipment and test servers; tester rack design, setup, qualification, maintenance, and capacity planning; debug of tester hardware, cabling, power, networking, PXE boot, and interface issues; hardware, firmware, software, and system integration; release of robust test infrastructure to production; equipment uptime and utilization improvement; preventive maintenance; calibration and spares management; vendor coordination; and development of scalable, cost-effective solutions for server manufacturing.
KEY RESPONSIBILITIES
Work closely with design, validation, test engineering, manufacturing, IT, quality, and operations teams to define and deploy tester infrastructure. Partner with equipment, rack, cable, fixture, and contract manufacturing vendors to ensure readiness, quality, and timely support. Lead tester equipment and test-server readiness from NPI through mass production, identifying risks that could limit capacity, yield, or product ramp. Build, configure, qualify, and sustain tester racks, test servers, switches, power distribution, cabling, fixtures, and diagnostic equipment. Troubleshoot equipment-level failures across server hardware, BIOS/BMC/firmware, Linux, networking, storage, power, thermal, and test interfaces. Drive tester uptime, utilization , repeatability, and capacity improvements through preventive maintenance, failure analysis, automation, and standardization. Establish equipment health monitoring, maintenance schedules, calibration controls, spare-parts strategy, and escalation processes. Develop and maintain tester setup procedures, qualification plans, troubleshooting guides, and technician training materials. Support server manufacturing bring-up, tester deployment, production recovery, and technical training onsite at the factory. Communicate equipment status, risks, root cause, containment, and recovery plans clearly to cross-functional stakeholders. Work independently in a fast-paced environment and provide technical leadership through ambiguous equipment and infrastructure issues. PREFERRED EXPERIENCES: At least 3 + years of relevant experience supporting manufacturing tester equipment, server test infrastructure, or production test operations. Hands-on experience with server hardware, tester racks, test fixtures, switches, storage, power distribution, cabling, and diagnostic equipment. Strong understanding of server architecture, including CPU/GPU, memory, PCIe, storage, networking, power, thermal, BIOS, BMC, and firmware. Experience configuring and administering Linux-based test servers and services such as DHCP, PXE, SSH, file transfer, and remote management. Familiarity with IPMI/Redfish, BIOS/BMC configuration, firmware updates, log collection, and server health monitoring. Proficiency in Linux shell scripting and/or Python for tester automation,