About this role
About the Role Join AMD’s Datacenter GPU Platform Application Engineering team as a System Application Engineer. You will collaborate with external OEM partners, internal development and validation teams, and cross-functional stakeholders to bring next-generation server platforms powered by AMD Instinct Accelerators to market and ensure their successful deployment in customer data centers. What You'll Do
- Manage technical interaction with OEM/ODM Partners to enable deployment of AMD Instinct Accelerators in Partner systems.
- Support Partners in the bring-up and validation of AMD Instinct GPUs; guide partners on use of AMD tools, qualification test methods, and analysis of test results.
- Lead the debugging of Partner/Customer issues (HW, firmware, driver), working with a cross-functional team and driving root cause investigations.
- Work with Partners on the development of manufacturing/screen tests to ensure reliability at scale.
- Understand Partner requirements and schedule, identify gaps in AMD offering and work with other stakeholders to close them.
- Author design guideline, technical presentations, and training material.
- Provide recommendations to improve customer experience with our SW and HW. What We're Looking For
- Solid years’ experience in Data Center system design, board design, validation, or Application Engineering, preferably in external customer facing roles.
- Strong knowledge in PC/server architecture and interfaces, experience with system level debug.
- Strong System Level debugging skills with hands-on experiences in system bring-up, HW debug, and performance optimizations on various system architectures.
- Understanding and experience working with Enterprise Linux environment (Ubuntu, CentOS/RHEL and SLES).
- Excellent oral and written communication skills to communicate technical results clearly and accurately.
- Test or design experience in system interconnect technologies (PCIe, XGMI, CXL, USB, I/O Controllers).
- Familiarity with various deployment models including cloud, virtualization and containers.
- Automation, orchestration, delivery via Kubernetes, Docker, or Mesos.
- Experience or knowledge of server firmware/BIOS settings, boot process, server monitoring and management SW.
- Experience relating to power and thermal management.
- Solid knowledge of Shell/BASH, C/C++, Python, or other framework.
- Experience with OpenCL, CUDA, or ROCm is a plus.
- Bilingual with both English and Chinese. Nice to Have
- MS or PhD is a plus. Compensation & Benefits
- Salary: USD 193,200 – USD 289,800 per year.
- Benefits offered: AMD benefits at a glance.