HPC Platform Engineer, Software, Center for Quantum Computing

Amazon Web Services (AWS)

San Francisco (CA)

On-site

USD 149,000 - 201,000

Full time

13 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Amazon Development Center U.S., Inc. in San Francisco seeks an HPC Platform Engineer to develop, automate, and maintain high-performance computing infrastructure on AWS for the MQS Center of Quantum Computing.

The role supports quantum computing hardware design and simulation with MPI-based parallelism and containerized environments. You will collaborate with scientists and engineers, translate complex science requirements into reliable, scalable infrastructure, and drive CI/CD pipelines,

Qualifications

  • Automate, deploy, and support large-scale infrastructure.
  • 2+ years designing or architecting scalable systems.
  • Experience with CI/CD pipelines.
  • Experience with Linux/Unix.
  • Experience using infrastructure-as-code to deploy cloud services.

Responsibilities

  • Administer and automate cloud-based HPC environments by deploying and maintaining clusters, managing OS and software stacks, building containers, provisioning users, and securing systems.
  • Support computational workflows by improving HPC robustness, performance, and availability across instance types and regions.
  • Anticipate and expand CQC computational capacity, leveraging AWS to accelerate the quantum hardware design cycle.
  • Design reproducible development environments and own dependency, build, and release management across interdependent research software projects.
  • Develop CI/CD pipelines and automate provisioning using infrastructure-as-code (e.g., Docker, AWS CDK).
  • Maintain Amazon's high security bar while enabling fast-paced development; implement observability and monitoring for rapid debugging and high uptime.

Skills

Automation
CI/CD
Python
Rust
Golang
C++
Java
Docker
Kubernetes
AWS

Tools

Docker
Kubernetes
AWS CDK
Terraform
CloudFormation

Job description

Description

The Models, Quantum, & Silicon (MQS) Center for Quantum Computing (CQC) is a multi-disciplinary team of scientists, engineers, and technicians, on a mission to develop a fault-tolerant quantum computer.


The Models, Quantum, & Silicon (MQS) Center for Quantum Computing (CQC) is a multi-disciplinary team of scientists, engineers, and technicians, on a mission to develop a fault-tolerant quantum computer.


We are looking to hire an HPC Platform Engineer to develop, automate, and maintain high-performance computing (HPC) infrastructure on AWS that CQC scientists and engineers use for quantum computing hardware design and simulation. You will work closely with our experimental and theoretical physics teams to enable large-scale HPC workloads with MPI-based parallelism on EC2 instances, manage graphical environments for computer-aided engineering applications, and accelerate the research computing lifecycle through automation and infrastructure-as-code. The ideal candidate will be able to translate high-level science and simulation requirements into reliable deliverables (including cluster orchestration, job scheduling, CI/CD pipelines, reproducible environments, and artifact management) that are performant, scalable, and secure. This requires someone who (1) has a strong desire to work within a team of scientists and engineers, (2) demonstrates ownership by initiating and driving projects to completion, and (3) is comfortable operating at the intersection of traditional HPC and cloud-native infrastructure.


Why MQS?


Models, Quantum, & Silicon is a brand new organization within Amazon. We're building the full stack of compute to power the next era of innovation. One Org. One Mission. Limitless Impact!


Diverse Experiences


Amazon values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying.


Work/Life Balance


We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.


Inclusive Team Culture


Here at MQS, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon (gender diversity) conferences, inspire us to never stop embracing our uniqueness.


Mentorship and Career Growth


We’re continuously raising our performance bar as we strive to become Earth’s Best Employer. That’s why you’ll find endless knowledge-sharing, mentorship and other career-advancing resources here to help you develop into a better-rounded professional.


Key job responsibilities

Responsibilities Include Some Combination Of The Following



  • Administer and automate cloud-based HPC environments by deploying and maintaining clusters, managing OS and software stacks, building containers, provisioning users, and securing systems.

  • Support computer-aided engineering and computational science workflows (e.g., Palace) by improving HPC environment robustness, performance, and availability across instance types and regions.

  • Anticipate and expand CQC computational capacity, leveraging AWS to accelerate the quantum hardware design cycle.

  • Design reproducible development environments and own dependency, build, and release management across interdependent research software projects.

  • Develop CI/CD pipelines and automate provisioning using infrastructure-as-code (e.g., Docker, AWS CDK).

  • Maintain Amazon's high security bar while enabling fast-paced development; implement observability and monitoring for rapid debugging and high uptime.


We are looking for candidates with strong engineering principles, a bias for action, superior problem-solving, and excellent communication skills. Working effectively within a team environment is essential. As an HPC Platform Engineer embedded in a research organization, you will have the opportunity to work on new ideas and stay abreast of the field of experimental quantum computation.


A day in the life

The majority of your time will be spent on projects that strengthen and scale our HPC platform and accelerate the computational workflows that underpin quantum hardware design. This requires working backwards from the needs of our science staff in the context of our larger experimental roadmap. You will translate science and software requirements into design proposals, balancing implementation complexity against time-to-delivery. Once a proposal has been reviewed and accepted, you'll drive implementation and coordinate with internal stakeholders to ensure a smooth rollout. You will work closely with engineers and scientists on the team to understand their simulation workloads, parallelism requirements, and tooling needs.


About The Team

You will be joining the Software team within the MQS Center of Quantum Computing. Our team is comprised of scientists and engineers who are building scalable software that enables quantum computing technologies.


Basic Qualifications


  • Experience in automating, deploying, and supporting large-scale infrastructure

  • Experience programming with at least one modern language such as Python, Ruby, Golang, Java, C++, C#, Rust

  • Experience with Linux/Unix

  • Experience with CI/CD pipelines build processes

  • 2+ years of designing or architecting (design patterns, reliability and scaling) of new and existing systems experience

  • Experience using infrastructure-as-code to design and deploy cloud services


Preferred Qualifications


  • Experience with distributed systems at scale

  • Experience in an AWS environment, including VPC, EC2, EBS, S3, SQS, CloudFormation and Lambda

  • Experience in network fundamentals (DNS, DHCP, TCP/IP, routing, switching, HTTP)

  • Experience in Kubernetes, Docker or containers ecosystem

  • Experience working with scientists in a research environment

  • Experience with high-performance computing (HPC) infrastructure


Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.


Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company’s reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.


Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.


Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.


The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.


USA, CA, San Francisco - 148,700.00 - 201,200.00 USD annually


Company - Amazon Development Center U.S., Inc.


Job ID: A10466636

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Design Automation Engineer, Software, Center for Quantum Computing
Senior Design Automation Engineer, Software, Center for Quantum Computing

Amazon Web Services (AWS) • San Francisco (CA)

On-site
USD 192,000 - 260,000
Health insurance
401(k) matching
Paid time off
+1
Quantum Applied Scientist, AWS Center for Quantum Computing
Quantum Applied Scientist, AWS Center for Quantum Computing

Amazon Science • Pasadena (CA)

On-site
USD 142,000 - 193,200
Quantum Applied Scientist, AWS Center for Quantum Computing
Quantum Applied Scientist, AWS Center for Quantum Computing

Amazon Web Services (AWS) • Pasadena (CA)

On-site
USD 143,000 - 193,000
System Development Engineer - Quantum Engineering, AWS CQC
System Development Engineer - Quantum Engineering, AWS CQC

Amazon Web Services (AWS) • San Francisco (CA)

On-site
USD 120,000 - 160,000
Applied Scientist, Quantum Algorithms, Center for Quantum Computing
Applied Scientist, Quantum Algorithms, Center for Quantum Computing

Amazon • Pasadena (CA)

On-site
USD 143,000 - 193,000
Health insurance
RSUs
Sign-on bonus
+1
Applied Scientist, Quantum Algorithms, Center for Quantum Computing
Applied Scientist, Quantum Algorithms, Center for Quantum Computing

Amazon • San Francisco (CA)

On-site
USD 143,000 - 222,000
Health insurance
RSUs
401(k) matching
+1
Quantum Hardware Engineer, Amazon's Center for Quantum Computing, Device Team
Quantum Hardware Engineer, Amazon's Center for Quantum Computing, Device Team

Socket.dev • Pasadena (CA)

On-site
USD 123,000 - 160,000
Quantum Applied Scientist, AWS Center for Quantum Computing
Quantum Applied Scientist, AWS Center for Quantum Computing

Amazon • Pasadena (CA)

On-site
USD 142,800 - 193,200
Health insurance
RSU equity
Paid time off
+1
Applied Scientist, Device Team, AWS Center for Quantum Computing
Applied Scientist, Device Team, AWS Center for Quantum Computing

Amazon Web Services (AWS) • Pasadena (CA)

On-site
USD 143,000 - 193,000
Health insurance
401(k) matching
Paid time off
+2
Applied Scientist, Quantum error correction team, Amazon Center for Quantum Computing
Applied Scientist, Quantum error correction team, Amazon Center for Quantum Computing

Amazon Web Services (AWS) • Pasadena (CA)

On-site
USD 143,000 - 193,000