Senior Software Engineer - SoC DevOps, MLA-MI - Annapurna Labs

Amazon Web Services (AWS)

Cupertino (CA)

On-site

USD 193,300 - 261,500

Full time

7 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Annapurna Labs (U.S.) Inc. in Cupertino, CA seeks a Senior SoC Software DevOps Engineer to own end-to-end CI/CD pipelines for SoC software, firmware, HAL, and modeling tools. You will enable rapid, reliable iteration across pre-silicon simulation and post-silicon deployments.

The role emphasizes automation observability and cross-environment releases, with a focus on reducing build times and accelerating tape-outs while maintaining strict quality standards.

Qualifications

  • Experience across the full software development life cycle including coding standards and operations.
  • Mentor or lead engineering teams.
  • Proficient in Python and at least one of Bash, Go, C++, or Java.

Responsibilities

  • Own end-to-end CI/CD pipelines and release processes for SoC software components.
  • Develop and maintain infrastructure automation and tooling for pre-silicon simulation and post-silicon deployments.
  • Build observability dashboards and alerting; improve development velocity and release quality.

Skills

Full SDLC
Mentor / Tech lead
Python
Bash
Go
C++
Java
Linux
Infrastructure-as-code
AWS services
Automation

Education

Bachelor's degree

Tools

Jenkins
Terraform
CloudFormation
AWS CDK

Job description

Description

The Senior SoC Software DevOps Engineer role centers on enabling the rapid and reliable development of software for AWSs most advanced custom machine learning chips. This position is critical to supporting the Trainium and Inferentia families of silicon which power large scale AI training at AWS. The engineer will serve as the primary owner of infrastructure that directly affects how quickly software teams can iterate on code for both pre silicon simulation environments and post silicon production deployments. By building robust automation and tooling the role ensures that tape outs for new chips stay on schedule and that software is ready to function immediately when first silicon becomes available. This work has a direct impact on AWSs ability to deliver advanced ML infrastructure to its largest customers.

Description

The Senior SoC Software DevOps Engineer role centers on enabling the rapid and reliable development of software for AWSs most advanced custom machine learning chips. This position is critical to supporting the Trainium and Inferentia families of silicon which power large scale AI training at AWS. The engineer will serve as the primary owner of infrastructure that directly affects how quickly software teams can iterate on code for both pre silicon simulation environments and post silicon production deployments. By building robust automation and tooling the role ensures that tape outs for new chips stay on schedule and that software is ready to function immediately when first silicon becomes available. This work has a direct impact on AWSs ability to deliver advanced ML infrastructure to its largest customers.

This role operates at the intersection of hardware and software requiring deep expertise in infrastructure engineering to solve unique challenges such as coordinating releases across isolated environments and validating firmware on real silicon. It is a foundational position for the SoC software teams as it frees engineers from infrastructure burdens allowing them to focus on feature development. Success in this role will be measured by improvements in development velocity release quality and the stability of systems that support multiple teams. The position demands a proactive approach to identifying bottlenecks and a strong ability to operate within novel technical contexts without prior domain knowledge in machine learning or chip design.

Key job responsibilities

The engineer will own the end to end CI/CD pipelines and release processes for all SoC software components including firmware hardware abstraction layers and modeling tools. This involves designing maintaining and evolving systems that produce reliable releases for both internal verification teams and external AWS services. A key task is ensuring these pipelines function across heterogeneous environments such as Corp networks and VPC. The role requires building qualification workflows that guarantee software meets strict quality standards before reaching customers or verification teams.

Another core duty is developing hardware in the loop test infrastructure that validates SoC software on actual silicon in laboratory and automated testing settings. This includes creating frameworks to run tests on real chips simulate pre silicon environments and integrate results into continuous integration workflows. Additionally the engineer must build observability tools such as dashboards that track build health test coverage and pipeline performance along with alerting systems that notify teams of regressions. A significant focus will be on identifying and removing friction in development workflows such as slow build times or complex release steps using data driven insights to prioritize improvements that accelerate team productivity. The role also involves solving novel problems like bridging disconnected environments and orchestrating synchronized releases across multiple domains.

About The Team

We're part of the SoC Software organization within Annapurna Labs (AWS). Our three software teams — uCode, HAL (Hardware Abstraction Layer), and Modeling — build the firmware, drivers, and virtual platforms for AWS's custom ML accelerator chips. We operate like a startup: small teams, high ownership, direct impact on AWS's most strategic silicon programs. This DevOps engineer will work across all three teams, with a mandate to improve velocity, quality, and developer experience for the entire SoC software organization.

Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship.

Basic Qualifications
  • 7+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
  • Experience as a mentor, tech lead or leading an engineering team
  • Experience programming in Python and at least one of: Bash, Go, C++, or Java
  • Experience with infrastructure-as-code (CDK, CloudFormation, Terraform, etc.)
  • Experience with AWS services (Lambda, S3, EC2, CloudWatch, IAM, Secrets Manager, etc.)
  • Experience with Linux-based build and development environments
Preferred Qualifications
  • Bachelor's degree in computer science or equivalent
  • Experience with Amazon's internal build and release systems (Brazil, Pipelines, CRUX, Apollo, etc.)
  • Experience building cross-environment or cross-account automation (e.g., bridging corporate and isolated/air-gapped environments)
  • Experience with Jenkins pipeline development and administration
  • Experience with hardware-in-the-loop testing or supporting hardware/silicon development teams
  • Experience building observability infrastructure: dashboards, metrics pipelines, alerting (CloudWatch, QuickSight, or similar)

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company’s reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&DD insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.

USA, CA, Cupertino - 193,300.00 - 261,500.00 USD annually

USA, TX, Austin - 168,100.00 - 227,400.00 USD annually

Company

Annapurna Labs (U.S.) Inc.

Job ID: A10385470

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer - SoC DevOps, MLA-MI - Annapurna Labs (AWS)
Senior Software Engineer - SoC DevOps, MLA-MI - Annapurna Labs (AWS)

Amazon • Austin (TX)

On-site
USD 168,100 - 227,400
RSUs
Health insurance
401(k) matching
+2
Senior SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS
Senior SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 193,000 - 262,000
Health insurance
401(k) matching
Paid time off
+2
SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS
SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 165,000 - 224,000
SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS
SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 143,000 - 195,000
Senior SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS
Senior SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS

Amazon • Cupertino (CA)

On-site
USD 193,000 - 261,500
RSUs
Health insurance
401(k) matching
+2
Software Development Engineer - Silicon Development Infrastructure
Software Development Engineer - Silicon Development Infrastructure

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 143,700 - 194,400
Health insurance
RSUs
SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS
SoC Systems Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS

Amazon • Cupertino (CA)

On-site
USD 165,200 - 223,600
RSUs
Health insurance
401(k) matching
+1
SDE, MLA hardware/software co-design, Annapurna Labs Machine Learning Acceleration
SDE, MLA hardware/software co-design, Annapurna Labs Machine Learning Acceleration

Amazon Web Services (AWS) • Cupertino (CA)

On-site
USD 157,000 - 213,000
Health insurance
401(k) matching
Paid time off
+1
Senior Virtual Platform Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS
Senior Virtual Platform Software Engineer, Annapurna Labs Machine Learning Accelerators, AWS

Amazon Web Services (AWS) • Austin (TX)

On-site
USD 100,000 - 228,000
Lead SDE C/C++ Hardware/Software Co-Design, Machine Learning Acceleration Systems (AWS)
Lead SDE C/C++ Hardware/Software Co-Design, Machine Learning Acceleration Systems (AWS)

Amazon • Cupertino (CA)

On-site
USD 193,300 - 261,500
Health insurance
401(k) matching
Paid time off
+1