Data Center Engineer

Advanced Micro Devices

Temple, Northern (TX, KY)

Hybrid

USD 120,000 - 180,000

Full time

34 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

AMD Benefits

Job summary

Advanced Micro Devices (AMD) is seeking a senior engineer to lead our datacenter graphics hardware/software subsystem projects for OEM partners and enterprise customers in Temple, TX. You will apply expertise in AI/ML, virtualization, and server GPU deployment to enable high-performance compute workloads.

In this role you’ll collaborate across teams, install and validate software on GPU/compute clusters, guide customers, and help drive hardware bring-up through end-of-life support.

Qualifications

  • Bachelor’s or Master’s in CS/EE or equivalent.
  • Experience with datacenter GPU/software stacks and virtualization.
  • Strong debugging, analytical and communication skills.

Responsibilities

  • Perform node and cluster-level software installation and validation for GPU/compute AI projects.
  • Resolve technical issues for customers using AMD Instinct products.
  • Provide technical guidance and support for server graphics and compute projects.
  • Build datacenter GPU containers for customer testing and deployment.
  • Qualify new software functionality for compatibility with customer requirements.
  • Collaborate with PMs to maintain schedules and deliverables.

Skills

Datacenter deployment
GPU/AI/ML systems
Linux administration
Troubleshooting
Networking
Project management

Education

Bachelor’s degree in Computer Engineering / Electrical Engineering / CS
Master’s degree (optional)

Tools

VMware
KVM
Docker
ROCm

Job description

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believetechnology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMDis shapingthefuture.

Whetheryou’redesigning next-gen processors, enabling AI breakthroughs, orbringing leading edge products to market, every role at AMD contributes to something bigger— technologythat moves the world forward.Join us and, together, we’ll advance your career.

THE ROLE:

In this critical and highly technical role, you will be responsible for executing AMD's Datacenter graphics hardware/software subsystem projects for AMD OEM partners and enterprise commercial end-customers. This position provides a unique opportunity to leverage your expertise in graphics, compute, datacenter technologies, virtualization, AI/Machine Learning, and program management to collaborate with customers utilizing AMD Instinct™ Accelerators.

You must be a team player with a strong commitment to meeting deadlines and the ability to thrive in a fast-paced, multi-tasking environment.

THE PERSON:

The ideal candidate possesses exceptional datacenter deployment and troubleshooting skills in AI GPU hardware, software, and networking. They are highly analytical, detail-oriented, self-motivated, and maintain a positive, results-driven attitude. This role may require up to 25% travel.

KEY RESPONSIBILITIES:
  • Perform node and cluster-level software installation and validation for GPU/compute AI and Machine Learning projects.
  • Resolve technical issues for customers utilizing AMD Instinct™ products.
  • Provide technical guidance and support to customers for server graphics and compute projects related to AI and Machine Learning workloads.
  • Build datacenter GPU dockers and containers for customer testing and deployment.
  • Qualify and assess new software functionality to ensure compatibility with customer requirements.
  • Assist development teams in identifying and resolving hardware/software technical issues throughout the product lifecycle, from initial hardware bring-up to end-of-life.
  • Follow procedures to communicate, report, and elevate incidents to AMD Management.
  • Collaborate with program managers to maintain project schedules, track action items, ensure deliverables are met, and provide project status updates to customers and AMD management.
  • Develop a strong understanding of the client’s business to ensure impactful and effective task completion.
PREFERRED EXPERIENCE:
  • Experience in:
  • Datacenter customer support roles.
  • Large-scale cluster deployment within hyperscale datacenters.
  • Server architecture and functionality, including remote management, network topologies, and graphics software/hardware subsystems.
  • Linux installation, setup, usage, and debugging.
  • Virtual environments (e.g., VMWare, Citrix, KVM, Microsoft) and virtual machine setup/management.
  • Datacenter GPU software stacks such as AMD ROCm™ or Nvidia CUDA.
  • Validating multimode AI clusters using AMD tools (e.g., AGFHC, RCCL RDMA) or equivalents.
  • AI/Machine Learning workloads, frameworks, and models.
  • Strong debugging, problem-solving, and analytical skills.
  • Excellent verbal and written communication skills for conveying technical information.
  • Self-starter with attention to detail, organizational skills, and the ability to multitask in a fast-paced environment.
ACADEMIC CREDENTIALS:
  • Bachelor’s or Master’s degree in computer engineering, Electrical Engineering, Computer Science or equivalent.
  • Technical certifications in relevant software systems are highly desirable
LOCATION:

Temple, TX

This role is not eligible for visa sponsorship.

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Systems Application Engineer
Systems Application Engineer

AMD • Austin (TX)

Hybrid
USD 120,000 - 180,000
Product Application Engineer- Data Center Deployment
Product Application Engineer- Data Center Deployment

AMD • California (MO)

On-site
USD 80,000 - 110,000
Health insurance
401(k) plan
Paid time off
Data Center Infrastructure Architect
Data Center Infrastructure Architect

CareerArc • North Carolina

On-site
USD 150,000 - 230,000
AI /HPC Data Center Lab Engineer
AI /HPC Data Center Lab Engineer

Advanced Micro Devices • Austin (TX)

On-site
USD 90,000 - 120,000
Product Development Engineer, Datacenter Systems (AI/HPC)
Product Development Engineer, Datacenter Systems (AI/HPC)

AMD • Austin (TX)

On-site
USD 110,000 - 170,000
AI & Datacenter Solutions Architect
AI & Datacenter Solutions Architect

AMD • South Carolina

On-site
USD 180,000 - 260,000
Principal Data Center GPU Performance Architect
Principal Data Center GPU Performance Architect

Advanced Micro Devices • Austin (TX)

On-site
USD 180,000 - 250,000
Technical Support Engineer
Technical Support Engineer

AMD • Austin (TX)

On-site
USD 120,000 - 180,000
Data Center Infrastructure Architect
Data Center Infrastructure Architect

Advanced Micro Devices • Northern (KY)

Hybrid
USD 150,000 - 210,000
Field Applications Engineer, Server Datacenter - Dell
Field Applications Engineer, Server Datacenter - Dell

AMD • Austin (TX)

On-site
USD 120,000 - 180,000