Data Center Engineer

AMD

United States

On-site

USD 120,000 - 170,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

AMD seeks a highly technical engineer to execute datacenter graphics hardware/software projects for OEM partners and enterprise customers. You will leverage expertise in graphics, compute, virtualization, and AI/ML to support AI workloads and collaborate with customers using AMD Instinct accelerators.

Ideal candidates have datacenter deployment experience, strong Linux skills, and the ability to work in a fast-paced environment with up to 25% travel.

Qualifications

  • Bachelor's or Master's in a related field.
  • Strong background in datacenter deployment and troubleshooting.
  • Experience with Linux, virtualization, and AI/ML workloads.

Responsibilities

  • Perform node and cluster-level software installation and validation for GPU/compute AI projects.
  • Resolve technical issues for customers using AMD Instinct products.
  • Provide guidance and support for server graphics and compute projects related to AI workloads.
  • Build datacenter GPU docks/containers for testing and deployment.
  • Assess new software functionality for customer compatibility.
  • Collaborate with PMs to maintain schedules and report status to customers.

Skills

Datacenter deployment
Troubleshooting
Linux
Cluster management
AI/ML workloads
Communication

Education

Bachelor's degree in Computer/Electrical Engineering or CS

Tools

AMD ROCm
NVIDIA CUDA
Docker
KVM

Job description

ADVANCE YOUR CAREER. ADVANCE THE WORLD.At AMD, we believe technology has the power to solve the world's most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future.Whether you're designing next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger - technology that moves the world forward. Join us and, together, we'll advance your career.

THE ROLE:

In this critical and highly technical role, you will be responsible for executing AMD's Datacenter graphics hardware/software subsystem projects for AMD OEM partners and enterprise commercial end-customers. This position provides a unique opportunity to leverage your expertise in graphics, compute, datacenter technologies, virtualization, AI/Machine Learning, and program management to collaborate with customers utilizing AMD Instinct™ Accelerators.You must be a team player with a strong commitment to meeting deadlines and the ability to thrive in a fast-paced, multi-tasking environment.

THE PERSON:

The ideal candidate possesses exceptional datacenter deployment and troubleshooting skills in AI GPU hardware, software, and networking. They are highly analytical, detail-oriented, self-motivated, and maintain a positive, results-driven attitude. This role may require up to 25% travel.

KEY RESPONSIBILITIES:
  • Perform node and cluster-level software installation and validation for GPU/compute AI and Machine Learning projects.
  • Resolve technical issues for customers utilizing AMD Instinct™ products.
  • Provide technical guidance and support to customers for server graphics and compute projects related to AI and Machine Learning workloads.
  • Build datacenter GPU dockers and containers for customer testing and deployment.
  • Qualify and assess new software functionality to ensure compatibility with customer requirements.
  • Assist development teams in identifying and resolving hardware/software technical issues throughout the product lifecycle, from initial hardware bring-up to end-of-life.
  • Follow procedures to communicate, report, and elevate incidents to AMD Management.
  • Collaborate with program managers to maintain project schedules, track action items, ensure deliverables are met, and provide project status updates to customers and AMD management.
  • Develop a strong understanding of the client's business to ensure impactful and effective task completion.
PREFERRED EXPERIENCE:

Experience in:

  • Datacenter customer support roles.
  • Large-scale cluster deployment within hyperscale datacenters.
  • Server architecture and functionality, including remote management, network topologies, and graphics software/hardware subsystems.
  • Linux installation, setup, usage, and debugging.
  • Virtual environments (e.g., VMWare, Citrix, KVM, Microsoft) and virtual machine setup/management.
  • Datacenter GPU software stacks such as AMD ROCm™ or Nvidia CUDA.
  • Validating multimode AI clusters using AMD tools (e.g., AGFHC, RCCL RDMA) or equivalents.
  • AI/Machine Learning workloads, frameworks, and models.
  • Strong debugging, problem-solving, and analytical skills.
  • Excellent verbal and written communication skills for conveying technical information.
  • Self-starter with attention to detail, organizational skills, and the ability to multitask in a fast-paced environment.
ACADEMIC CREDENTIALS:

Bachelor's or Master's degree in computer engineering, Electrical Engineering, Computer Science or equivalent.

LOCATION:

Temple, TX

This role is not eligible for visa sponsorship.

Benefits offered are described: AMD benefits at a glance .

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants' needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD's "Responsible AI Policy" is available here.

This posting is for an existing vacancy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Center Engineer
Data Center Engineer

Advanced Micro Devices • Temple (TX), Northern (KY)

Hybrid
USD 120,000 - 180,000
AMD Benefits
Data Center Engineer
Data Center Engineer

Socket.dev • Town of Texas (WI)

On-site
USD 140,000 - 190,000
Data Center Engineer
Data Center Engineer

AMD • Temple (TX)

On-site
USD 95,000 - 140,000
Senior Datacenter Platform/Debug Engineer
Senior Datacenter Platform/Debug Engineer

Advanced Micro Devices, Inc. • Austin (TX)

On-site
USD 110,000 - 160,000
AMD benefits at a glance
Data Center Infrastructure Architect
Data Center Infrastructure Architect

CareerArc • North Carolina

On-site
USD 150,000 - 230,000
Product Application Engineer- Data Center Deployment
Product Application Engineer- Data Center Deployment

AMD • California (MO)

On-site
USD 80,000 - 110,000
Health insurance
401(k) plan
Paid time off
AI /HPC Data Center Lab Engineer
AI /HPC Data Center Lab Engineer

Advanced Micro Devices • Austin (TX)

On-site
USD 90,000 - 120,000
Systems Application Engineer
Systems Application Engineer

AMD • Austin (TX)

Hybrid
USD 120,000 - 180,000
Product Development Engineer, Datacenter Systems (AI/HPC)
Product Development Engineer, Datacenter Systems (AI/HPC)

AMD • Austin (TX)

On-site
USD 110,000 - 170,000
Field Applications Engineer, Server Datacenter - Dell
Field Applications Engineer, Server Datacenter - Dell

AMD • Austin (TX)

On-site
USD 120,000 - 180,000