Principal Solutions Engineering – AI server/rack Infrastructure
Listed on 2026-09-07
-
IT/Tech
Systems Engineer, Hardware Engineer -
Engineering
Systems Engineer, Hardware Engineer
AMD Data Center Platform Engineering Group
At AMD, we believe technology has the power to solve the world's most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future.
Whether you're designing next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger — technology that moves the world forward. Join us and, together, we'll advance your career.
The RoleAMD is searching for a dynamic and experienced Principal Member of Technical Staff to own system design support, rack-level bring-up, and critical customer engagement for our cutting-edge AMD Instinct™ product line. In this high-visibility role, you will act as the technical bridge between AMD's internal system architects, platform development teams, and our OEM partners. You will not only influence the design and architecture of AI solutions but also lead hands-on debug and validation efforts at customer locations.
As a technical leader, you will drive engineering, root cause analysis, and influence future roadmaps based on field execution.
- System Architecture & Design Support
- Solution Optimization:
Partner deeply with customers to architect and optimize Rack-Scale AI solution deployments using AMD Instinct GPUs. - Design Reviews:
Provide support of design reviews for customer platform/rack designs; proactively flag areas for modification to improve quality, performance and competitive advantage. - Bring-Up, Debug & Validation
- Documentation & Best Practices:
Deliver comprehensive technical documentation, best practices, and reference architectures to streamline the adoption and deployment of AMD AI platforms. - Hands-on Engineering:
Drive hands-on rack, platform, and component-level debug and validation. This includes complex stress testing, issue reproductions, and deep-dive root cause analysis. - Issue Resolution:
Lead customer issue resolution efforts, gathering diagnostics, managing critical escalations, and driving long-term process improvements to ensure customer success. - System Firmware Debug & Deployment:
Lead debug efforts for system firmware (BIOS, BMC) during initial bring-up and large-scale deployment phases. Ensure seamless integration between hardware, firmware, and software stacks, and resolve interaction issues in customer environments. - End-Customer Debug & Sustaining:
Own the technical support interface for end customers, provide high-level engineering for deployed fleets. - Leadership & Strategy
- Cross-Functional Alignment:
Represent debug progress, technical insights, and status with clarity and impact at the leadership level, ensuring alignment and accountability across cross-functional teams. - Roadmap Influence:
Provide regular, detailed technical feedback from the field to directly influence AMD's software and hardware roadmaps. - Future Architecture:
Drive future product architecture decisions by leveraging unique insights gained from deep customer execution engagement. - Mentorship:
Build a culture of ownership, accountability, and technical excellence within the team, while actively mentoring senior engineers and emerging technical leaders.
- Advanced experience in system architecture, hardware/firmware debug, and customer-facing engineering roles (HPC or AI/ML focus preferred).
- Deep understanding of Server/Rack system architecture (x86, GPU, PCIe, Interconnects).
- Strong proficiency in System Firmware (BIOS/UEFI, BMC/OpenBMC) debug, update flows, and deployment strategies.
- Experience with system bring-up and debugging tools (oscilloscopes, logic analyzers, ITP, JTAG).
- Knowledge of power delivery, thermal management, and mechanical form factors in datacenter environments.
- Leadership:
Proven track record of leading technical teams through complex problem-solving scenarios and interacting with executive leadership. - Travel:
Ability to travel to customer, factory and company locations
Bachelors, Masters, or PhD in Electrical Engineering, Computer Engineering, or Computer Science.
Locations- Seattle, WA., Austin, TX., or Santa Clara, CA.
This role does not support visa sponsorship.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).