Sr. GPU/Accelerator Hardware Development Engineer, Annapurna Labs

Amazon Web Services (AWS)

Austin (TX)

On-site

USD 159,000 - 215,000

Full time

25 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Annapurna Labs (U.S.) Inc. seeks a Lead Hardware Design Engineer to own the system design, validation, and hardware integration across the AWS data center fleet.

You will collaborate with silicon, firmware, and system software teams to enable optimal co-design, delivering world-class performance, quality, and cost improvements. You will lead end-to-end server hardware development, including PCB/motherboard design, carrier boards, and high-speed interconnects, while maintaining high standards in

Qualifications

  • BS or MS in Electrical or Computer Engineering or equivalent.
  • 5+ years in high-speed system design and validation.
  • Experience with schematic and layout tools.
  • Experience driving HW development and testing through production.
  • Strong knowledge of power, signal integrity, and analog/digital circuits.
  • Experience with lab equipment and test instrumentation.

Responsibilities

  • Lead system design, validation, and integration of hardware in AWS data centers.
  • Collaborate with cross-functional teams to improve quality, reliability, and cost.
  • Drive hardware development lifecycle from concept to production.
  • Ensure on-time delivery and performance optimization for large-scale server deployments.

Skills

High-Speed system design
Project leadership
Electrical fundamentals

Education

Electrical or Computer Engineering degree

Tools

Schematic tools
Layout tools
Lab equipment
Production flow experience

Job description

Description

Would you like to develop the Next Generation of AI accelerator compute systems? Lead bleeding-edge HW development projects? Have you heard of Amazon Web Services (AWS) Project Rainer? This is the opportunity to be a part of a fast-moving innovation team that is changing the world of AI at massive scale. At AWS Trainium we develop a complete vertical stack system, from our own Silicon to Hardware to Software and deploy directly to our customers in our own Data Centers


We are seeking experienced Lead System Design Engineers to build the next generation of our cloud server infrastructure, Project Rainier. Project Rainier is a massive $11 billion Amazon Web Services (AWS) AI infrastructure initiative, featuring one of the world's largest compute clusters dedicated to training and running Anthropic’s Claude AI models. It utilizes over 500,000 custom Trainium2 chips, designed for high-performance AI training.


As a member of the AWS Trainium Machine Learning Acceleration team you’ll be responsible for the System design and optimization of hardware in our data centers. You’ll provide leadership in the application of new technologies to large scale server deployments in a continuous effort to deliver a world-class customer experience. This is a fast-paced, intellectually challenging position, and you’ll work with thought leaders in multiple technology areas. You’ll have high standards for yourself and everyone you work with, and you’ll be constantly looking for ways to improve your products performance, quality and cost. We’re changing industry, and we want individuals who are ready for this challenge and want to reach beyond what is possible today.


Key job responsibilities


We are looking for a Lead Hardware Design Engineer with strong skills in both hardware and software. In this role, you will be responsible for system design, validation, and integration of hardware in the AWS fleet through its entire life cycle. You will work cross functionally with AWS monitoring teams, members of the Hardware Design team, and additional teams across AWS to improve quality and reliability of products operating in the fleet.


We are looking for candidates who thrive in a fast-paced start-up like environment and work independently to deliver multiple projects in parallel. To be successful, you need to be highly motivated and detailed oriented while meeting the highest standards and time to market, cost and quality goals.


Basic Qualifications


  • BS or MS degree in Electrical or Computer Engineering (EE / CE)

  • Minimum of 5 years of experience with High-Speed system design and validation

  • Experience with Schematic and layout tools.

  • Drive ODM HW development and testing and be part of the Production flow definition team

  • Strong knowledge in electrical engineering fundamentals, power & signal integrity, and analog/digital circuits

  • Able to drive component selection and validation of electrical, mechanical components, cables

  • Experience with hardware development process and system development across full product life cycles

  • Experience using lab equipment such as bench power supplies, high-speed oscilloscopes, logic analyzers, spectrum analyzers, VNA’s, and thermal chambers

  • Experience with supply chain management


Preferred Qualifications


  • Lead end-to-end server hardware development lifecycle from Concept, Architecture, Design, Validation and Production

  • Drive PCB board design for server motherboards, accelerator carrier boards, and high-speed interconnect boards.

  • Collaborate with silicon, firmware, and system software teams to enable optimal hardware/software co-design.

  • Improve compute density, power efficiency, and network bandwidth utilization.

  • Drive root cause analysis for hardware issues during validation and production.


Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.


Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company’s reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.


Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.


USA, CA, Cupertino - 183,000.00 - 247,600.00 USD annually


USA, TX, Austin - 159,200.00 - 215,300.00 USD annually


USA, WA, Seattle - 159,200.00 - 215,300.00 USD annually


Company - Annapurna Labs (U.S.) Inc.

Job ID: A10427058

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. GPU/Accelerator Hardware Development Engineer, Annapurna Labs
Sr. GPU/Accelerator Hardware Development Engineer, Annapurna Labs

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 159,000 - 215,000
Sr. GPU/Accelerator Hardware Development Engineer
Sr. GPU/Accelerator Hardware Development Engineer

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 183,000 - 248,000
Health insurance
401(k) matching
Paid time off
+1
Sr. GPU/Accelerator Hardware Development Engineer (AWS)
Sr. GPU/Accelerator Hardware Development Engineer (AWS)

Amazon • Austin (TX)

On-site
USD 159,000 - 215,000
Health insurance
401(k) matching
Paid time off
+1
Sr. GPU/Accelerator Hardware Development Engineer, Annapurna Labs
Sr. GPU/Accelerator Hardware Development Engineer, Annapurna Labs

Amazon • Cupertino (CA)

On-site
USD 183,000 - 248,000
Health insurance
401(k) matching
Paid time off
+1
Hardware Engineer - ML Acceleration, Annapurna Labs
Hardware Engineer - ML Acceleration, Annapurna Labs

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 136,000 - 184,000
Systems Development Engineer (AWS Generative AI & ML Servers), AWS HW Engineering
Systems Development Engineer (AWS Generative AI & ML Servers), AWS HW Engineering

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 149,000 - 201,000
Software Engineer II, Annapurna Labs ML Acceleration System Software
Software Engineer II, Annapurna Labs ML Acceleration System Software

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 144,000 - 194,000
Systems Development Engineer (AWS Generative AI & ML Servers), AWS HW Engineering
Systems Development Engineer (AWS Generative AI & ML Servers), AWS HW Engineering

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 129,000 - 175,000
Health insurance
401(k) matching
Paid time off
+1
Systems Development Engineer (AWS Generative AI & ML Servers), AWS HW Engineering
Systems Development Engineer (AWS Generative AI & ML Servers), AWS HW Engineering

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 129,000 - 175,000
Senior SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS
Senior SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
401(k) matching
Paid time off
+1