HPC Platform Engineer, Software, Center for Quantum Computing

Amazon

San Francisco, Northern (CA, KY)

Hybrid

USD 149,000 - 201,000

Full time

7 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Amazon's MQS Center for Quantum Computing seeks an HPC Platform Engineer to develop, automate, and maintain HPC infrastructure on AWS that supports quantum hardware design and simulation.

You will enable large-scale HPC workloads with MPI on EC2, manage graphical environments for CAE apps, and deliver reproducible environments, CI/CD pipelines, and secure, scalable systems.

Qualifications

  • Experience automating, deploying, and supporting large-scale infrastructure.
  • Proficient in one modern language such as Python, Ruby, Golang, Java, C++, C#, Rust.
  • Experience with Linux/Unix.
  • Experience with CI/CD pipelines build processes.
  • 2+ years of designing or architecting distributed systems.
  • Experience using infrastructure-as-code to design and deploy cloud services.
  • Experience with distributed systems at scale.
  • Experience in an AWS environment, including VPC, EC2, EBS, S3, SQS, CloudFormation and Lambda.
  • Experience with network fundamentals (DNS, DHCP, TCP/IP, routing, HTTP).
  • Experience with Kubernetes, Docker or containers ecosystem.
  • Experience working with scientists in a research environment.
  • Experience with high-performance computing (HPC) infrastructure.

Responsibilities

  • Administer and automate cloud-based HPC environments by deploying and maintaining clusters, managing OS and software stacks, building containers, provisioning users, and securing systems.
  • Support computer-aided engineering and computational science workflows (e.g., Palace) by improving HPC environment robustness, performance, and availability across instance types and regions.
  • Anticipate and expand CQC computational capacity, leveraging AWS to accelerate the quantum hardware design cycle.
  • Design reproducible development environments and own dependency, build, and release management across interdependent research software projects.
  • Develop CI/CD pipelines and automate provisioning using infrastructure-as-code (e.g., Docker, AWS CDK).
  • Maintain Amazon's high security bar while enabling fast-paced development; implement observability and monitoring for rapid debugging and high uptime.

Skills

Python
Ruby
Golang
Java
C++
C#
Rust
Linux/Unix
CI/CD pipelines
Kubernetes
AWS
Docker

Tools

Docker
Kubernetes
AWS CDK
CloudFormation
Terraform

Job description

HPC Platform Engineer, Software, Center for Quantum Computing

ID lavoro: 10466636 | Amazon Development Center U.S., Inc.


The Models, Quantum, & Silicon (MQS) Center for Quantum Computing (CQC) is a multi-disciplinary team of scientists, engineers, and technicians, on a mission to develop a fault-tolerant quantum computer. We are looking to hire an HPC Platform Engineer to develop, automate, and maintain high-performance computing (HPC) infrastructure on AWS that CQC scientists and engineers use for quantum computing hardware design and simulation. You will work closely with our experimental and theoretical physics teams to enable large-scale HPC workloads with MPI-based parallelism on EC2 instances, manage graphical environments for computer-aided engineering applications, and accelerate the research computing lifecycle through automation and infrastructure-as-code. The ideal candidate will be able to translate high-level science and simulation requirements into reliable deliverables (including cluster orchestration, job scheduling, CI/CD pipelines, reproducible environments, and artifact management) that are performant, scalable, and secure. This requires someone who has a strong desire to work within a team of scientists and engineers, demonstrates ownership by initiating and driving projects to completion, and is comfortable operating at the intersection of traditional HPC and cloud-native infrastructure.


Why MQS?

Models, Quantum, & Silicon is a brand new organization within Amazon. We're building the full stack of compute to power the next era of innovation. One Org. One Mission. Limitless Impact!


Diverse Experiences

Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying.


Work/Life Balance

We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.


Inclusive Team Culture

Here at MQS, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (gender diversity) conferences, inspire us to never stop embracing our uniqueness.


Mentorship and Career Growth

We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.


Key job responsibilities


  • Administer and automate cloud-based HPC environments by deploying and maintaining clusters, managing OS and software stacks, building containers, provisioning users, and securing systems.

  • Support computer-aided engineering and computational science workflows (e.g., Palace) by improving HPC environment robustness, performance, and availability across instance types and regions.

  • Anticipate and expand CQC computational capacity, leveraging AWS to accelerate the quantum hardware design cycle.

  • Design reproducible development environments and own dependency, build, and release management across interdependent research software projects.

  • Develop CI/CD pipelines and automate provisioning using infrastructure-as-code (e.g., Docker, AWS CDK).

  • Maintain Amazon's high security bar while enabling fast-paced development; implement observability and monitoring for rapid debugging and high uptime.


We are looking for candidates with strong engineering principles, a bias for action, superior problem-solving, and excellent communication skills. Working effectively within a team environment is essential. As an HPC Platform Engineer embedded in a research organization, you will have the opportunity to work on new ideas and stay abreast of the field of experimental quantum computation.


A day in the life

The majority of your time will be spent on projects that strengthen and scale our HPC platform and accelerate the computational workflows that underpin quantum hardware design. This requires working backwards from the needs of our science staff in the context of our larger experimental roadmap. You will translate science and software requirements into design proposals, balancing implementation complexity against time-to-delivery. Once a proposal has been reviewed and accepted, you'll drive implementation and coordinate with internal stakeholders to ensure a smooth rollout. You will work closely with engineers and scientists on the team to understand their simulation workloads, parallelism requirements, and tooling needs.


About the team

You will be joining the Software team within the MQS Center of Quantum Computing. Our team is comprised of scientists and engineers who are building scalable software that enables quantum computing technologies.



  • Experience in automating, deploying, and supporting large-scale infrastructure

  • Experience programming with at least one modern language such as Python, Ruby, Golang, Java, C++, C#, Rust

  • Experience with Linux/Unix

  • Experience with CI/CD pipelines build processes

  • 2+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience

  • Experience using infrastructure-as-code to design and deploy cloud services

  • Experience with distributed systems at scale

  • Experience in an AWS environment, including VPC, EC2, EBS, S3, SQS, CloudFormation and Lambda

  • Experience in network fundamentals (DNS, DHCP, TCP/IP, routing, switching, HTTP)

  • Experience with Kubernetes, Docker or containers ecosystem

  • Experience working with scientists in a research environment

  • Experience with high-performance computing (HPC) infrastructure


Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.


Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company’s reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.


Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.


Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.


Qualifiche Preferenziali


  • Experience with distributed systems at scale

  • Experience in an AWS environment, including VPC, EC2, EBS, S3, SQS, CloudFormation and Lambda

  • Experience in network fundamentals (DNS, DHCP, TCP/IP, routing, switching, HTTP)

  • Experience with Kubernetes, Docker or containers ecosystem

  • Experience working with scientists in a research environment

  • Experience with high-performance computing (HPC) infrastructure


Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.


Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company’s reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.


Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.


Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.


The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits .


USA, CA, San Francisco - 148,700.00 - 201,200.00 USD annually

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

HPC Platform Engineer, Software, Center for Quantum Computing
HPC Platform Engineer, Software, Center for Quantum Computing

Amazon Web Services (AWS) • San Francisco (CA)

On-site
USD 149,000 - 201,000
Research Software Engineer, Calibration, MQS Center for Quantum Computing
Research Software Engineer, Calibration, MQS Center for Quantum Computing

Amazon • San Francisco (CA)

On-site
USD 172,000 - 222,000
Senior Applied Scientist, QEC SW, MQS Center for Quantum Computing
Senior Applied Scientist, QEC SW, MQS Center for Quantum Computing

Amazon • Seattle (WA), Northern (KY)

On-site
USD 192,000 - 260,000
Health insurance
401(k) matching
Paid time off
+1
Senior Applied Scientist, QEC SW, MQS Center for Quantum Computing
Senior Applied Scientist, QEC SW, MQS Center for Quantum Computing

Amazon • San Francisco (CA)

On-site
USD 192,000 - 260,000
FPGA Engineer, MQS Center for Quantum Computing
FPGA Engineer, MQS Center for Quantum Computing

Amazon Web Services (AWS) • Pasadena (CA)

On-site
USD 144,000 - 194,000
Amazon Quantum Applied Science Intern - Quantum Technologies team · 4 locationsBoston, MA, SF +2 Quantum Applied Science Intern - Quantum Technologies team 4 locationsBoston, MA, SF +2 5d ago
Amazon Quantum Applied Science Intern - Quantum Technologies team · 4 locationsBoston, MA, SF +2 Quantum Applied Science Intern - Quantum Technologies team 4 locationsBoston, MA, SF +2 5d ago

Amazon Inc. • Pasadena (CA), Northern (KY)

Hybrid
USD 115,000 - 156,000
Applied Scientist, Quantum Algorithms, Center for Quantum Computing
Applied Scientist, Quantum Algorithms, Center for Quantum Computing

Amazon • San Francisco (CA)

On-site
USD 172,000 - 222,000
Health insurance
Stock options
Paid time off
+1
Applied Scientist, Quantum Algorithms, Center for Quantum Computing
Applied Scientist, Quantum Algorithms, Center for Quantum Computing

Amazon • Pasadena (CA)

On-site
USD 143,000 - 193,000
Health insurance
RSUs
Sign-on bonus
+1
Quantum Research Engineer, Quantum Computing
Quantum Research Engineer, Quantum Computing

Amazon Science • Pasadena (CA)

On-site
USD 143,000 - 193,000
Senior Applied Scientist, Superconducting Digital and Cryogenic Control Electronics,Center for Quantum Computing, Physical Design and Simulation, Enabling Technologies (AWS)
Senior Applied Scientist, Superconducting Digital and Cryogenic Control Electronics,Center for Quantum Computing, Physical Design and Simulation, Enabling Technologies (AWS)

Amazon • San Francisco (CA)

On-site
USD 192,000 - 260,000