Principal Cloud & AI Workload Performance Analysis Engineer

Advanced Micro Devices

Town of Texas (WI)

On-site

USD 140,000 - 210,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

AMD is seeking a systems-minded performance engineer in Austin, TX to own workloads exposed to hyperscaler CPUs, AI systems and accelerator utilization. You will measure cloud-native performance, price-performance and performance-per-watt, comparing AMD EPYC against Intel and Arm platforms.

You will move from containers and orchestration to NUMA and memory bandwidth, explaining business impact with concise, evidence-based results. Collaboration with architecture teams is expected.

Qualifications

  • Hands-on benchmarking performance of cloud, distributed or accelerated systems.
  • Strong Linux systems skills and scripting for deployment/analysis.
  • Experience with containers, orchestration and cloud-native workloads at scale.
  • Ability to design fair experiments across different CPU architectures.
  • Familiarity with hardware performance counters and system telemetry.

Responsibilities

  • Define cloud-native and AI workload taxonomy for analysis and forecasting.
  • Design reproducible methods for web services, containers, and scale-out data processing.
  • Compare AMD/Intel/Arm/cloud environments with aligned software versions.
  • Analyze host-CPU impact on accelerator utilization and end-to-end throughput.
  • Model performance-per-watt, density and price-performance.
  • Track factors affecting cross-ISA deployability.
  • Explain bottlenecks with architecture experts.
  • Forecast scaling trends and assess risks.
  • Develop cross-ISA methodology documentation.
  • Support validation with partners and produce technical reports.

Skills

Benchmarking
Linux systems
Containers & orchestration
AI workloads
CPU-accelerator data

Education

Bachelors or Masters in electrical or computer engineering

Tools

Automation scripts
Profiling tools
Performance counters

Job description

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believetechnology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMDis shapingthefuture.

Whetheryou’redesigning next-gen processors, enabling AI breakthroughs, orbringing leading edge products to market, every role at AMD contributes to something bigger— technologythat moves the world forward.Join us and, together, we’ll advance your career.

THE ROLE

AMD is looking for a systems-minded performance engineer to own the workloads most exposed to hyperscaler custom CPUs, Arm ecosystem momentum and AI-system architecture. You will measure and explain cloud-native performance, price-performance, performance-per-watt, software maturity and host-CPU effects on accelerator utilization. The work requires fair methods for comparing AMD EPYC with Intel, Arm Neoverse-based platforms, cloud-custom CPUs, NVIDIA Grace-class systems and other emerging designs.

THE PERSON

You are curious at every layer of the system. You can move from containers, orchestration and application throughput down to NUMA placement, memory bandwidth, interconnect behavior and CPU-to-accelerator data movement - then explain the business consequence in a concise, evidence-based way. You care as much about deployability and software maturity as about a benchmark score.

KEY RESPONSIBILITIES
  • Define and maintain the cloud-native and AI workload taxonomy used in competitive analysis and forecasting.
  • Design reproducible methods for web services, microservices, containers, scale-out data processing, caching, cloud infrastructure and CPU-support functions in AI systems.
  • Compare AMD, Intel, Arm and cloud-custom environments using aligned software versions, tuning policies, instance shapes and service-level objectives.
  • Analyze host-CPU impact on accelerator utilization, input pipelines, communication overheads, memory movement, NUMA behavior and end-to-end AI system throughput.
  • Measure and model performance-per-watt, density, utilization and price-performance where the inputs can be normalized defensibly.
  • Track compiler, kernel, library, orchestration and migration factors that affect Arm and cross-ISA deployability.
  • Work with architecture experts to explain memory, interconnect and platform bottlenecks behind observed results.
  • Translate current scaling behavior and software trends into 24-36 month forecast inputs and early risk or advantage assessments.
  • Develop cross-ISA methodology documentation that can withstand partner, customer and internal technical scrutiny.
  • Support validation with approved cloud, OEM, ISV and ecosystem partners and produce decision-ready technical and executive reports.
KEY RESPONSIBILITIES
  • Deep hands-on experience benchmarking and profiling cloud, distributed or accelerated-system workloads.
  • Strong Linux systems skills and proficiency with automation or scripting for deployment and analysis.
  • Experience with containers, orchestration and cloud-native workloads at meaningful scale.
  • Ability to design fair experiments across dissimilar CPU architectures and platform configurations.
  • Experience with application profiling, hardware performance counters and system telemetry.
  • Understanding of AI host-side bottlenecks, CPU-to-accelerator data movement, NUMA, memory bandwidth and interconnect effects.
  • Strong technical writing and presentation skills, plus the ability to collaborate with external partners and reproduce results outside one lab.
PREFERRED RESPONSIBILITIES
  • Hands-on experience with Arm Neoverse or cloud-custom Arm platforms.
  • Experience benchmarking major cloud-service-provider instances and services.
  • Experience measuring AI system throughput or accelerator utilization as a function of host-CPU and platform behavior.
  • Familiarity with LPDDR or HBM-attached CPU systems, coherent CPU-accelerator links, high-speed networking or storage pipelines.
  • Participation in MLCommons, OCP or other relevant benchmark or standards communities.
WHY THIS OPPORTUNITY STANDS OUT
  • Work at the intersection of cloud-native software, custom silicon and next-generation AI systems.
  • Own full-stack investigations from application behavior to architecture and rack-level economics.
  • Influence optimization and product discussions before competitive narratives are set.
  • Build relationships and technical credibility across AMD, cloud providers and ecosystem partners.
ACADEMIC CREDENTIALS
  • Bachelors or Masters degree in electrical or computer engineering
LOCATION

Austin, Texas

This role is not eligible for visa sponsorship.

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

This posting is for an existing vacancy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Cloud & AI Workload Performance Analysis Engineer
Principal Cloud & AI Workload Performance Analysis Engineer

Socket.dev • Town of Texas (WI)

Hybrid
USD 140,000 - 210,000
Principal Cloud & AI Workload Performance Analysis Engineer
Principal Cloud & AI Workload Performance Analysis Engineer

AMD • Austin (TX)

Hybrid
USD 120,000 - 190,000
Hybrid work model
Principal Cloud & AI Workload Performance Analysis Engineer
Principal Cloud & AI Workload Performance Analysis Engineer

Advanced Micro Devices, Inc. • Austin (TX)

Hybrid
USD 140,000 - 170,000
Benefits at a glance
Chief Competitive Performance & Microarchitecture Analysis
Chief Competitive Performance & Microarchitecture Analysis

Advanced Micro Devices • Austin (TX)

On-site
USD 180,000 - 260,000
Chief Competitive Performance & Microarchitecture Analysis
Chief Competitive Performance & Microarchitecture Analysis

AMD • Austin (TX)

Hybrid
USD 180,000 - 240,000
AMD Benefits
Chief Competitive Performance & Microarchitecture Analysis
Chief Competitive Performance & Microarchitecture Analysis

AMD • Austin (TX)

On-site
USD 230,000 - 320,000
Senior Manager, AI Power & Performance Engineering
Senior Manager, AI Power & Performance Engineering

AMD • United States

Hybrid
USD 180,000 - 240,000
Performance Architect - x86 & Customer Solutions
Performance Architect - x86 & Customer Solutions

Advanced Micro Devices • Austin (TX)

On-site
USD 120,000 - 190,000
AMD benefits at a glance
Health insurance
Senior Manager, AI Power & Performance Engineering
Senior Manager, AI Power & Performance Engineering

Advanced Micro Devices • Austin (TX)

Hybrid
USD 150,000 - 190,000
AMD benefits
Principal Datacenter GPU Performance Architect
Principal Datacenter GPU Performance Architect

AMD • Austin (TX)

On-site
USD 180,000 - 240,000