Senior Field Application Engineer - OEM
NVIDIA (Eightfold)
- Location
- US, TX, Austin; US, TX, Remote
- Work model
- Remote
- Level
- Senior
- Posted
- Sep 18, 2026
Skills
About this role
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world. Are you ready to become part of a group that’s redefining computer graphics, PC gaming, and accelerated computing? At NVIDIA, we’re unlocking the boundless possibilities of AI to pioneer the next chapter in computing. As a Senior Field Application Engineer-OEM, you’ll hold a meaningful role in evaluating the feasibility of new builds and managing new qualifications and certifications. This role is distinguished by the rare opportunity to work hands-on with groundbreaking technology and elite engineering teams to develop inventive solutions! Join us in Round Rock, TX, and help us accomplish what has never been done before!
What you'll be doing
Serve as the technical lead for sustaining engineering related to NVIDIA data center products used by OEM customers. Manage complex issues from initial triage and containment to root-cause analysis, corrective action, and resolution. Perform system- and board-level failure analysis to isolate hardware, firmware, software, thermal, power, and integration-related problems. Facilitate the complete RMA process, including failure-data collection, return authorization, material tracking, engineering disposition, and communication of findings. Reproduce field failures in laboratory environments, analyze failure trends, and drive preventive improvements in product quality and serviceability. Collaborate with OEMs, ODMs, manufacturing partners, and NVIDIA engineering, quality, reliability, operations, and supply-chain teams to resolve critical issues. Provide on-site technical support and deliver clear failure-analysis reports, corrective-action updates, troubleshooting procedures, and executive-level communications. What we need to see: BS or MS in Electrical Engineering, Computer Engineering, Computer Science, or a related discipline—or equivalent experience. 5+ years of relevant experience in Field Application Engineering, sustaining engineering, failure analysis, systems engineering, or product engineering. Hands-on experience debugging complex server or computing platforms at both the system and board levels. Demonstrated experience managing field failures or RMA cases through root-cause analysis and corrective-action closure using methods such as 8D, FRACAS, or fault-tree analysis. Strong understanding of server architecture, including CPUs, GPUs, memory, storage, power delivery, cooling, and high-speed interconnects. Knowledge of PCIe and experience troubleshooting BMC/IPMI or Redfish, BIOS/UEFI, Linux, drivers, firmware, system logs, and hardware telemetry. Experience with hardware bring-up, board validation, system qualification, and L1–L10 server integration and manufacturing test stages. Ability to interpret schematics, block diagrams, diagnostic data, manufacturing test results, and thermal or electrical measurements. Excellent project-management and communication skills. Able to prioritize tasks and work effectively with business and engineering teams. Must operate from the Round Rock area and travel as required. Ways to stand out from the crowd: Direct experience supporting NVIDIA HGX, DGX, or MGX platforms or comparable GPU-accelerated data center systems. Experience owning sustaining engineering, failure-analysis, or RMA programs for an OEM, ODM, cloud service