Senior Lab Systems Engineer

Advanced Micro Devices

Austin (TX)

On-site

USD 120,000 - 160,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Advanced Micro Devices (AMD) is seeking a hands-on Platform Systems Engineer for its Datacenter Platform Engineering Group (DPEG). You will troubleshoot complex hardware, firmware, and software issues across GPU and server platforms, collaborating with hardware, firmware, software, and validation teams to ensure reliable operation of large-scale compute environments.

The ideal candidate enjoys solving challenging technical problems in fast-paced datacenter environments, mentoring others, and

Qualifications

  • Bachelor's or Master's degree in a technical discipline.
  • Experience with system-level hardware, firmware, and software debugging.
  • Datacenter/server/HPC/AI infrastructure experience preferred.
  • Strong root-cause analysis and triage skills.
  • Excellent communication and collaboration abilities.
  • Ability to mentor junior engineers.
  • Documentation and organizational skills.
  • Experience with Linux and scripting (Python/Bash).

Responsibilities

  • Support datacenter deployments and maintain uptime of large-scale compute systems.
  • Perform system-level debugging across hardware, firmware, software, and OS layers.
  • Investigate and resolve complex platform issues impacting GPU/server infrastructure.
  • Support system bring-up, initialization, validation, and operational readiness.
  • Provide technical leadership and guidance to junior engineers.
  • Document debugging methodologies and best practices.
  • Collaborate with cross-functional teams to drive resolution and improvements.

Skills

Troubleshooting
Root cause analysis
Communication
Mentoring
Documentation
Cross-functional collaboration
Analytical skills
Problem solving

Education

Bachelor's or Master's degree in Computer Engineering or related

Tools

Linux
Python
Bash
Debug tools

Job description

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believetechnology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shapingthefuture.

Whetheryou’redesigning next-gen processors, enabling AI breakthroughs, orbringing leading edge products to market, every role at AMD contributes to something bigger— technologythat moves the world forward.Join us and, together, we’ll advance your career.

THE ROLE:

Join AMD's Datacenter Platform Engineering Group (DPEG) and help support the deployment, availability, and operational success of next-generation AI and HPC infrastructure. As a Platform Systems Engineer, you will work on cutting-edge GPU and server platforms, partnering with hardware, firmware, software, validation, and datacenter engineering teams to troubleshoot complex system issues and ensure reliable operation of large-scale compute environments.

This role offers the opportunity to work directly with advanced datacenter technologies, participate in system bring-up and deployment activities, and become a key contributor in resolving critical platform-level issues. Ideal candidates enjoy solving challenging technical problems, collaborating across multiple engineering disciplines, and making a direct impact on the success of AMD's datacenter infrastructure.

THE PERSON:

The ideal candidate is a hands‑on systems engineer who enjoys deep technical troubleshooting and thrives in fast‑paced datacenter environments. They possess strong analytical skills, can quickly isolate and resolve complex issues, and are comfortable working across hardware, firmware, and software layers of a system.

Successful candidates will demonstrate:

  • Strong troubleshooting and root cause analysis skills
  • Excellent communication and collaboration abilities
  • A proactive and self‑driven approach to problem solving
  • Ability to mentor and guide junior engineers
  • Strong documentation and organizational skills
  • Comfort operating in highly technical and mission‑critical environments
  • A passion for learning new technologies and solving complex engineering challenges
KEY RESPONSIBILITIES:
  • Support datacenter deployments and help maintain the availability and uptime of large‑scale compute systems.
  • Perform system‑level debugging and triage across hardware, firmware, software, and operating system layers.
  • Investigate and resolve complex platform issues impacting GPU and server infrastructure.
  • Support system bring‑up, initialization, validation, and operational readiness activities.
  • Utilize industry‑standard debug tools and diagnostic methods to identify root causes.
  • Provide technical leadership and guidance to junior engineers during troubleshooting activities.
  • Document debug methodologies, troubleshooting procedures, and best practices.
  • Collaborate with cross‑functional engineering teams to drive issue resolution and continuous improvement.
PREFERRED EXPERIENCE:
  • System‑level hardware, firmware, and software debugging experience
  • Datacenter, server, HPC, or AI infrastructure environments
  • Root cause analysis and triage of complex platform issues
  • GPU, PCIe, memory, retimer, networking, and system architecture knowledge
  • RAS (Reliability, Availability, Serviceability) concepts and methodologies
  • Linux operating system experience
  • Python, Bash, or similar scripting experience
  • Hands‑on experience with industry‑standard debug tools and diagnostics
  • Server bring‑up, system initialization, and validation activities
  • Technical leadership, mentoring, and cross‑functional collaboration
ACADEMIC CREDENTIALS:
  • Bachelor's or Master’s degree preferred in Computer Engineering, Electrical Engineering, Computer Science, or a related technical discipline
LOCATION:

Rockdale, Texas | 100% Onsite

THIS ROLE IS NOT ELIGIBLE FOR VISA SUPPORT
LI-CS1

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee‑based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third‑party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Datacenter Platform/Debug Engineer
Senior Datacenter Platform/Debug Engineer

Advanced Micro Devices, Inc. • Austin (TX)

On-site
USD 110,000 - 160,000
AMD benefits at a glance
Senior Lab Systems Engineer
Senior Lab Systems Engineer

AMD • Rockdale (TX)

On-site
USD 110,000 - 160,000
Senior Lab Systems Engineer
Senior Lab Systems Engineer

Socket.dev • Austin (TX)

On-site
USD 120,000 - 180,000
Senior Network Systems Engineer
Senior Network Systems Engineer

Socket.dev • Austin (TX)

On-site
USD 90,000 - 120,000
AMD benefits
Senior Network Systems Engineer
Senior Network Systems Engineer

AMD • Rockdale (TX)

On-site
USD 90,000 - 130,000
Senior Technologist
Senior Technologist

Advanced Micro Devices, Inc. • Secaucus (NJ)

On-site
USD 55,000 - 85,000
Benefits
Senior Technologist
Senior Technologist

AMD • Secaucus (NJ)

On-site
USD 65,000 - 95,000
Sr. Field Applications Engineer, Datacenter & AI Systems Debug and Deployment Support
Sr. Field Applications Engineer, Datacenter & AI Systems Debug and Deployment Support

Advanced Micro Devices • Austin (TX)

On-site
USD 90,000 - 130,000
Principal Solutions Engineering – AI server/rack Infrastructure
Principal Solutions Engineering – AI server/rack Infrastructure

Advanced Micro Devices • Seattle (WA)

On-site
USD 180,000 - 240,000
AI /HPC Data Center Lab Engineer
AI /HPC Data Center Lab Engineer

Advanced Micro Devices • Austin (TX)

On-site
USD 90,000 - 120,000