Principal Solutions Engineering – AI server/rack Infrastructure

Advanced Micro Devices

Seattle (WA)

On-site

USD 180,000 - 240,000

Full time

47 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Advanced Micro Devices, Seattle-based, seeks a Principal Member of Technical Staff to own system design support, rack‑level bring‑up, and critical customer engagement for AMD Instinct products.

You will bridge internal architects with OEM partners, drive field debugging, and influence future roadmap through hands‑on leadership and expert engineering across hardware, firmware, and software. This role operates across onsite customer sites and labs.

Qualifications

  • Advanced experience in system architecture, hardware/firmware debug, and customer‑facing engineering roles (HPC or AI/ML focus preferred).
  • Deep understanding of Server/Rack system architecture (x86, GPU, PCIe, Interconnects).
  • Strong proficiency in System Firmware (BIOS/UEFI, BMC/OpenBMC) debug, update flows, and deployment strategies.
  • Experience with system bring‑up and debugging tools (oscilloscopes, logic analyzers, ITP, JTAG).
  • Knowledge of power delivery, thermal management, and mechanical form factors in datacenter environments.
  • Leadership: Proven track record of leading technical teams through complex problem‑solving scenarios and interacting with executive leadership.
  • Travel: Ability to travel to customer, factory and company locations

Responsibilities

  • System Architecture & Design Support: own rack‑level bring‑up and design reviews with customers.
  • Solution Optimization: architect and optimize Rack‑Scale AI deployments using AMD Instinct GPUs.
  • Bring‑Up, Debug & Validation: deliver hands‑on debugging, stress testing, and root‑cause analysis.
  • Documentation & Best Practices: create reference architectures and deployment guides for AMD AI platforms.
  • End‑Customer Debug & Sustaining: provide high‑level engineering for deployed fleets and respond to escalations.
  • Leadership & Strategy: influence roadmaps and mentor senior engineers; align with cross‑functional teams.

Skills

System architecture
Hardware/ Firmware debug
Customer-facing engineering
Leadership

Education

Bachelors/Masters/PhD in ECE/CE/CS

Tools

Oscilloscopes
Logic analyzers
JTAG
OpenBMC/BIOS debugging

Job description

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believe technology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future.

Whether you’re redesigning next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger— technology that moves the world forward. Join us and, together, we’ll advance your career.

THE ROLE:

AMD's Data Center Platform Engineering Group (DPEG) is designing, developing, and delivering innovative technology infrastructure enabling the digital world. We create cloud-enabling server/rack solutions that help the world’s leading companies turn their ideas into reality. Our customers are future-focused and so are we, always a step ahead of the next challenge. As experts in engineering, manufacturing, and supply chain, we’re the bridge between problem and solution for the world’s leading OEM & ODM partners and cloud services providers. Our customers depend on us to solve their most complex server/rack design needs. Come and join our Data Center Platform Engineering Group where we are building amazing, powered products with amazing people.

THE PERSON:

AMD is searching for a dynamic and experienced Principal Member of Technical Staff to own system design support, rack-level bring‑up, and critical customer engagement for our cutting‑edge AMD Instinct™ product line. In this high-visibility role, you will act as the technical bridge between AMD’s internal system architects, platform development teams, and our OEM partners. You will not only influence the design and architecture of AI solutions but also lead hands‑on debug and validation efforts at customer locations. As a technical leader, you will drive engineering, root cause analysis, and influence future roadmap based on field execution.

KEY RESPONSIBILITIES:
  • System Architecture & Design Support
  • Solution Optimization: Partner deeply with customers to architect and optimize Rack-Scale AI solution deployments using AMD Instinct GPUs.
  • Design Reviews: Provide support of design review for customer platform/rack designs; proactively flag areas for modification to improve quality, performance and competitive advantage.
  • Bring‑Up, Debug & Validation
  • Documentation & Best Practices: Deliver comprehensive technical documentation, best practices, and reference architectures to streamline the adoption and deployment of AMD AI platforms.
  • Hands‑on Engineering: Drive hands‑on rack, platform, and component‑level debug and validation. This includes complex stress testing, issue reproductions, and deep‑dive root cause analysis.
  • Issue Resolution: Lead customer issue resolution efforts, gathering diagnostics, managing critical escalations, and driving long‑term process improvements to ensure customer success.
  • System Firmware Debug & Deployment: Lead debug efforts for system firmware (BIOS, BMC) during initial bring‑up and large‑scale deployment phases. Ensure seamless integration between hardware, firmware, and software stacks, and resolve interaction issues in customer environments.
  • End‑Customer Debug & Sustaining: Own the technical support interface for end customers, provide high‑level engineering for deployed fleets.
  • Leadership & Strategy
  • Cross‑Functional Alignment: Represent debug progress, technical insights, and status with clarity and impact at the leadership level, ensuring alignment and accountability across cross‑functional teams.
  • Roadmap Influence: Provide regular, detailed technical feedback from the field to directly influence AMD’s software and hardware roadmaps.
  • Future Architecture: Drive future product architecture decisions by leveraging unique insights gained from deep customer execution engagement.
  • Mentorship: Build a culture of ownership, accountability, and technical excellence within the team, while actively mentoring senior engineers and emerging technical leaders.
PREFERRED EXPERIENCE:
  • Advanced experience in system architecture, hardware/firmware debug, and customer‑facing engineering roles (HPC or AI/ML focus preferred).
  • Deep understanding of Server/Rack system architecture (x86, GPU, PCIe, Interconnects).
  • Strong proficiency in System Firmware (BIOS/UEFI, BMC/OpenBMC) debug, update flows, and deployment strategies.
  • Experience with system bring‑up and debugging tools (oscilloscopes, logic analyzers, ITP, JTAG).
  • Knowledge of power delivery, thermal management, and mechanical form factors in datacenter environments.
  • Leadership: Proven track record of leading technical teams through complex problem‑solving scenarios and interacting with executive leadership.
  • Travel: Ability to travel to customer, factory and company locations
ACADEMIC CREDENTIALS:

Bachelors, Masters, or PhD in Electrical Engineering, Computer Engineering, or Computer Science.

LOCATIONS:
  • Seattle, WA., Austin, TX., or Santa Clara, CA.

This role does not support visa sponsorship.

#LI-CB1

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third‑party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Solutions Engineering – AI server/rack Infrastructure
Principal Solutions Engineering – AI server/rack Infrastructure

AMD • Seattle (WA)

On-site
USD 180,000 - 260,000
Principal Server and AI Product Architect
Principal Server and AI Product Architect

AMD • United States

Hybrid
USD 180,000 - 240,000
Benefits at a glance
Principal Enterprise AI/HPC GPU Systems Architect
Principal Enterprise AI/HPC GPU Systems Architect

Advanced Micro Devices • Austin (TX)

On-site
USD 210,000 - 260,000
AI Instinct System Management Architect
AI Instinct System Management Architect

AMD • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Lead Systems Management Architect
Lead Systems Management Architect

Socket.dev • Austin (TX)

On-site
USD 170,000 - 250,000
Software Solutions Architect
Software Solutions Architect

AMD • Austin (TX)

On-site
USD 130,000 - 160,000
Comprehensive benefits package
Lead Systems Management Architect
Lead Systems Management Architect

Advanced Micro Devices • Austin (TX)

On-site
USD 180,000 - 260,000
Principal Server and AI Product Architect
Principal Server and AI Product Architect

Advanced Micro Devices • Austin (TX)

On-site
USD 180,000 - 240,000
AMD benefits
Principal Enterprise AI/HPC GPU Systems Architect
Principal Enterprise AI/HPC GPU Systems Architect

AMD • Austin (TX)

On-site
USD 180,000 - 240,000
AMD benefits at a glance
AI Instinct System Management Architect
AI Instinct System Management Architect

Advanced Micro Devices • Santa Clara (CA), Northern (KY)

On-site
USD 180,000 - 240,000