Cloud Hardware Development Engineer, AWS Hardware Engineering Services, Specialized Platforms a[...]

Amazon

Cupertino (CA)

On-site

USD 157,000 - 213,000

Full time

11 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Health insurance
401(k) matching
Parental leave
Restricted stock units (RSUs)

Job summary

Amazon Web Services (AWS) HWEngS CHDE role in Cupertino, CA, focuses on fleet health, predictive detection, and auto-remediation for enterprise server racks. You will own diagnostics across hardware and software layers and drive zero-touch operations in a high-demand data-center context.

You’ll work with EC2 and DC Ops teams to scale the fleet, improve performance and reduce costs, while mentoring others and delivering robust server solutions for AWS infrastructure.

Qualifications

  • Bachelor's degree or equivalent in electrical/computer engineering or related field.
  • 3+ years designing servers or complex products.
  • Experience in design verification plans and functional tests.
  • Root cause analysis across firmware, kernel, driver, thermal, power, and hardware layers.

Responsibilities

  • Scale the operation of a massive fleet and lead server integration and delivery.
  • Develop automated monitoring, failure analysis, and remediation services.
  • Collaborate with AWS software teams to tailor server solutions for the AWS environment.
  • Support launch of servers into production and operate the fleet.

Skills

Scripting
Debugging
Linux

Education

Bachelor's degree in electrical engineering / computer engineering or equivalent
3+ years of relevant technical position experience

Tools

Linux commands

Job description

Cloud Hardware Development Engineer, AWS Hardware Engineering Services, Specialized Platforms and Servers

Job ID: 10516878 | Amazon Web Services, Inc.

Amazon Web Services (AWS) Hardware Engineering Services (HWEngS) owns the new product development (NPI) and operation of all AWS global infrastructure. In other words, we’re the people who keep the cloud running. We support all AWS data centers and all of the servers, storage, networking, power, and cooling equipment that ensure our customers have continual access to the innovation they rely on. We work on the most challenging problems, with thousands of variables impacting the supply chain, and we’re looking for talented people who want to help. You’ll join a diverse team of software, hardware, and network engineers, supply chain specialists, security experts, operations managers, and other vital roles. You’ll collaborate with people across AWS to help us deliver the highest standards for safety and security while providing seemingly infinite capacity at the lowest possible cost for our customers, and you’ll experience an inclusive culture that welcomes bold ideas and empowers you to own them to completion.

Cloud Hardware Development Engineer (CHDE): The Amazon Web Services (AWS) Hardware Engineering Services (HWEngS) Specialized Platforms and Servers team creates Enterprise rack solutions for Amazon’s innovative web services. We are seeking experienced CHDEs to own the fleet health, diagnostics, and automation of the Enterprise rack solutions:

  • Designing and implementing predictive failure detection systems using telemetry, sensor data, error trends, and log correlation to identify hardware issues before they cause a customer impact.
  • Driving toward zero-touch operations by building detection, diagnostics, and remediation of faults without human intervention
  • Debugging complex system failures in time-sensitive settings personally diving deep when the problem demands it.
  • Completing root cause analysis correlating across firmware, kernel, driver, thermal, power, and physical layers.

What you will do: As a member of the Specialized Platforms and Servers team, you’ll be responsible for collaborating with Elastic Cloud Compute (EC2) service teams and Data Center Operations to maintain fleet health in all the locations we have servers.

You will work closely with internal teams, suppliers, and external partners capturing lessons learned while operating the fleet to ensure next generation designs are of the highest quality, constantly looking for ways to improve your product performance, quality and cost.

Key job responsibilities

As a CHDE you will be responsible for scaling how we operate our massive existing & rapidly growing fleet. You will lead the integration and delivery of servers, support the development of automated monitoring, and failure analysis services to operate, debug, and scale our servers. You will work closely with other AWS software teams to tailor and operate servers solutions for the AWS environment. You will support launching our servers into production and operating our fleet of servers.

A day in the life

Your day to day responsibilities will be solving operational challenges to our existing fleet with the goal of improving the current customer experience as well as developing improved systems for future designs.

About the team

The team is comprised of CHDE's, System Development Engineers and Technical Program Managers, all with the common goal of delivering the best specialized server fleet possible to our customers.

Why AWS

Amazon Web Services (AWS) is the world’s most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating — that’s why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.

Inclusive Team Culture

Here at AWS, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (diversity) conferences, inspire us to never stop embracing our uniqueness.

Work/Life Balance

We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.

Mentorship and Career Growth

We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.

Diverse Experiences

Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying.

Basic Qualifications
  • Bachelor's degree in electrical engineering, computer engineering, or equivalent, or 3+ years of relevant technical position experience.
  • 3+ years of experience in server level design for compute or other complex product design.
  • Experience in developing design verification plans and functional test procedures.
  • Experience in board and server root cause analysis and resolution.
Preferred Qualifications
  • 5+ years of experience in complex product development such as servers, network switches, or other highly integrated devices with hardware, software, and service aspects.
  • Proficient scripting, debug abilities, and Linux operations commands.
  • Experience deploying and operating hardware and applications across large data centers.
  • Meets/exceeds Amazon’s leadership principles requirements for this role
  • Meets/exceeds Amazon’s functional/technical depth and complexity for this role

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company’s reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits .

  • USA, CA, Cupertino - 157,300.00 - 212,800.00 USD annually
  • USA, WA, Seattle - 136,000.00 - 184,000.00 USD annually
Important FAQs for current Government employees

Before proceeding, please review the following FAQs

https://www.amazon.jobs/en/faqs#faqs-for-us-government-employees

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Cloud Hardware Development Engineer, AWS Hardware Engineering Services, Specialized Platforms a[...]
Cloud Hardware Development Engineer, AWS Hardware Engineering Services, Specialized Platforms a[...]

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 157,000 - 213,000
Sign-on payments
Restricted stock units (RSUs)
Health insurance
+2
Cloud Hardware Development Engineer, AWS Hardware Engineering Services, Specialized Platforms a[...]
Cloud Hardware Development Engineer, AWS Hardware Engineering Services, Specialized Platforms a[...]

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 136,000 - 184,000
Health insurance
401(k) matching
Paid time off
Sr. System Development Engineer, AWS Mainstream Compute
Sr. System Development Engineer, AWS Mainstream Compute

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 174,000 - 235,000
Health insurance
401(k) matching
paid time off
+1
Hardware Development Engineer II, AWS Mainstream Compute
Hardware Development Engineer II, AWS Mainstream Compute

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 136,000 - 184,000
Hardware Development Engineer II, AWS Mainstream Compute
Hardware Development Engineer II, AWS Mainstream Compute

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 136,000 - 184,000
System Development Engineer, AWS Mainstream Compute
System Development Engineer, AWS Mainstream Compute

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 149,000 - 201,000
System Development Engineer, AWS Mainstream Compute
System Development Engineer, AWS Mainstream Compute

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 129,000 - 175,000
Sr. Technical Program Manager, Hardware NPI, Edge & High Performance Accelerator Servers for AI/ML
Sr. Technical Program Manager, Hardware NPI, Edge & High Performance Accelerator Servers for AI/ML

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 171,000 - 231,000
Sr. Technical Program Manager, Hardware NPI, Edge & High Performance Accelerator Servers for AI/ML
Sr. Technical Program Manager, Hardware NPI, Edge & High Performance Accelerator Servers for AI/ML

Amazon Web Services (AWS) • Seattle (WA)

On-site
USD 149,000 - 201,000
Sr. Technical Program Manager, Hardware NPI, Edge & High Performance Accelerator Servers for AI/ML
Sr. Technical Program Manager, Hardware NPI, Edge & High Performance Accelerator Servers for AI/ML

Amazon • Austin (TX)

On-site
USD 149,000 - 201,000