Principal Cloud & AI Workload Performance Analysis Engineer

Advanced Micro Devices, Inc.

Austin (TX)

Hybrid

USD 140,000 - 170,000

Full time

7 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Benefits at a glance

Job summary

Advanced Micro Devices, Inc. (AMD) is seeking a systems-minded performance engineer to own workloads exposed to hyperscaler CPUs and AI systems. You will measure cloud-native performance, price-performance, and throughput, comparing AMD EPYC with Intel, Arm, and other architectures.

Strong emphasis on memory, interconnect, and data movement at scale. You will move from containers and orchestration down to NUMA placement and CPU-accelerator interactions, translating results into business impact

Qualifications

  • Bachelor's or Master's in electrical or computer engineering.
  • Experience benchmarking cloud-native and AI workloads.
  • Strong Linux skills and scripting for deployment/analysis.
  • Experience with containers, orchestration, and cloud workloads at meaningful scale.
  • Ability to design fair experiments across dissimilar CPU architectures and platforms.

Responsibilities

  • Define and maintain cloud-native and AI workload taxonomy for competitive analysis and forecasting.
  • Design reproducible methods for web services, containers, scale-out data processing, and AI system support.
  • Compare AMD, Intel, Arm and cloud environments with aligned software versions and SLAs.
  • Analyze host-CPU impact on accelerator utilization, memory movement and end-to-end AI throughput.
  • Measure and model performance-per-watt, density, and price-performance across platforms.
  • Track compiler, kernel, library, and orchestration factors affecting cross-ISA deployability.
  • Explain bottlenecks with architecture experts and translate trends into forecast inputs.
  • Develop cross-ISA methodology documentation for partners and customers.
  • Support validation with cloud/OEM/ISV partners and produce exec-ready reports.

Skills

Systems thinking
Analytical reasoning
Technical writing
Collaboration

Education

Bachelor's or Master's in Electrical or Computer Engineering

Tools

Linux
Docker
Kubernetes
Performance profiling tools
CPU architectures knowledge

Job description

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believetechnology has the power to solve the world's most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMDis shapingthefuture.

Whetheryou'redesigning next-gen processors, enabling AI breakthroughs, orbringing leading edge products to market, every role at AMD contributes to something bigger- technologythat moves the world forward.Join us and, together, we’ll advance your career.

THE ROLE:

AMD is looking for a systems-minded performance engineer to own the workloads most exposed to hyperscaler custom CPUs, Arm ecosystem momentum and AI-system architecture. You will measure and explain cloud-native performance, price-performance, performance-per-watt, software maturity and host-CPU effects on accelerator utilization. The work requires fair methods for comparing AMD EPYC with Intel, Arm Neoverse-based platforms, cloud-custom CPUs, NVIDIA Grace-class systems and other emerging designs.

THE PERSON:

You are curious at every layer of the system. You can move from containers, orchestration and application throughput down to NUMA placement, memory bandwidth, interconnect behavior and CPU-to-accelerator data movement - then explain the business consequence in a concise, evidence-based way. You care as much about deployability and software maturity as about a benchmark score.

KEY RESPONSIBILITIES:
  • Define and maintain the cloud-native and AI workload taxonomy used in competitive analysis and forecasting.
  • Design reproducible methods for web services, microservices, containers, scale-out data processing, caching, cloud infrastructure and CPU-support functions in AI systems.
  • Compare AMD, Intel, Arm and cloud-custom environments using aligned software versions, tuning policies, instance shapes and service-level objectives.
  • Analyze host-CPU impact on accelerator utilization, input pipelines, communication overheads, memory movement, NUMA behavior and end-to-end AI system throughput.
  • Measure and model performance-per-watt, density, utilization and price-performance where the inputs can be normalized defensibly.
  • Track compiler, kernel, library, orchestration and migration factors that affect Arm and cross-ISA deployability.
  • Work with architecture experts to explain memory, interconnect and platform bottlenecks behind observed results.
  • Translate current scaling behavior and software trends into 24-36 month forecast inputs and early risk or advantage assessments.
  • Develop cross-ISA methodology documentation that can withstand partner, customer and internal technical scrutiny.
  • Support validation with approved cloud, OEM, ISV and ecosystem partners and produce decision-ready technical and executive reports.
KEY RESPONSIBILITIES:
  • Deep hands-on experience benchmarking and profiling cloud, distributed or accelerated-system workloads.
  • Strong Linux systems skills and proficiency with automation or scripting for deployment and analysis.
  • Experience with containers, orchestration and cloud-native workloads at meaningful scale.
  • Ability to design fair experiments across dissimilar CPU architectures and platform configurations.
  • Experience with application profiling, hardware performance counters and system telemetry.
  • Understanding of AI host-side bottlenecks, CPU-to-accelerator data movement, NUMA, memory bandwidth and interconnect effects.
  • Strong technical writing and presentation skills, plus the ability to collaborate with external partners and reproduce results outside one lab.
PREFERRED RESPONSIBILITIES:
  • Hands-on experience with Arm Neoverse or cloud-custom Arm platforms.
  • Experience benchmarking major cloud-service-provider instances and services.
  • Experience measuring AI system throughput or accelerator utilization as a function of host-CPU and platform behavior.
  • Familiarity with LPDDR or HBM-attached CPU systems, coherent CPU-accelerator links, high-speed networking or storage pipelines.
  • Participation in MLCommons, OCP or other relevant benchmark or standards communities.
WHY THIS OPPORTUNITY STANDS OUT:
  • Work at the intersection of cloud-native software, custom silicon and next-generation AI systems.
  • Own full-stack investigations from application behavior to architecture and rack-level economics.
  • Influence optimization and product discussions before competitive narratives are set.
  • Build relationships and technical credibility across AMD, cloud providers and ecosystem partners.
ACADEMIC CREDENTIALS:
  • Bachelors or Masters degree in electrical or computer engineering
LOCATION:

Austin, Texas

This role is not eligible for visa sponsorship.

#LI-RW1

#LI-Hybrid

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants' needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD's "Responsible AI Policy" is available here.

This posting is for an existing vacancy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Cloud & AI Workload Performance Analysis Engineer
Principal Cloud & AI Workload Performance Analysis Engineer

AMD • Austin (TX)

Hybrid
USD 120,000 - 190,000
Hybrid work model
Principal Cloud & AI Workload Performance Analysis Engineer
Principal Cloud & AI Workload Performance Analysis Engineer

Socket.dev • Town of Texas (WI)

Hybrid
USD 140,000 - 210,000
Principal Cloud & AI Workload Performance Analysis Engineer
Principal Cloud & AI Workload Performance Analysis Engineer

Advanced Micro Devices • Town of Texas (WI)

On-site
USD 140,000 - 210,000
Chief Competitive Performance & Microarchitecture Analysis
Chief Competitive Performance & Microarchitecture Analysis

Advanced Micro Devices • Austin (TX)

On-site
USD 180,000 - 260,000
Chief Competitive Performance & Microarchitecture Analysis
Chief Competitive Performance & Microarchitecture Analysis

AMD • Austin (TX)

Hybrid
USD 180,000 - 240,000
AMD Benefits
Chief Competitive Performance & Microarchitecture Analysis
Chief Competitive Performance & Microarchitecture Analysis

AMD • Austin (TX)

On-site
USD 230,000 - 320,000
Senior Manager, AI Power & Performance Engineering
Senior Manager, AI Power & Performance Engineering

AMD • United States

Hybrid
USD 180,000 - 240,000
Senior Manager, AI Power & Performance Engineering
Senior Manager, AI Power & Performance Engineering

Advanced Micro Devices • Austin (TX)

Hybrid
USD 150,000 - 190,000
AMD benefits
Performance Architect - x86 & Customer Solutions
Performance Architect - x86 & Customer Solutions

Advanced Micro Devices • Austin (TX)

On-site
USD 120,000 - 190,000
AMD benefits at a glance
Health insurance
Principal Datacenter GPU Performance Architect
Principal Datacenter GPU Performance Architect

Advanced Micro Devices, Inc. • Austin (TX)

On-site
USD 150,000 - 190,000