Senior Site Reliability Engineer

Anduril Industries

Washington (Washington County)

On-site

USD 140,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Full Family Health Coverage
16 Weeks Paid Parental Leave for All-C
Family Planning & Support
Incentivized Time Off
Mental Health Resources
Financial Planning
Unlimited Provisions
Professional Development

Job summary

Anduril Industries is seeking a Site Reliability Engineer (SRE) to join our Irvine-based team. You will lead development of Kubernetes cloud infrastructure, DevOps, CI/CD, and tooling to improve deployment speed and reliability.

Responsibilities include managing AWS/Azure/on-prem deployments, architecting scalable infrastructure, and driving best practices for resilience and high availability across large-scale environments.

Qualifications

  • Networking, cloud technologies, applications development or cybersecurity experience.
  • 6+ years of engineering experience.
  • Deep knowledge of the Kubernetes ecosystem (Docker, Helm, ArgoCD, Terraform).
  • Experience with Go, Python, Rust, or C++.
  • Eligible to obtain and maintain a U.S. Secret security clearance.
  • Computer Science degree or equivalent.

Responsibilities

  • Lead the development of Kubernetes cloud infrastructure and DevOps practices.
  • Architect, deploy and maintain infrastructure with cloud providers and Kubernetes (EKS).
  • Develop and maintain CI/CD pipelines for automated deployment.
  • Promote SRE best practices in resilience, performance monitoring, and HA.
  • Manage cloud deployments in AWS, Azure, and on-premise environments.
  • Collaborate with multi-disciplinary teams on internal and external deployments.
  • Design, develop, and deliver IaC solutions with Terraform and Python.
  • Improve operational capabilities of core product through tooling for large-scale deployments.
  • Lead organization in scalable, sustainable deployment delivery at scale.
  • Perform root cause analysis and drive continuous improvement.

Skills

Networking
Cloud technologies
Kubernetes
Go
Python
Rust
C++
Data-driven analysis
Security clearance
Linux

Education

Computer Science degree or equivalent

Tools

Docker
Terraform
ArgoCD
qemu
KubeVirt
Linux

Job description

Responsibilities
  • We are looking for a Site Reliability Engineer (SRE) to join AGD, our rapidly growing team in Irvine
  • SREs work with external stakeholders to determine the technical direction of cloud deployments and deliver with speed through analysis, design and code
  • They are comfortable leading large, focused projects
  • They lead in the development of Kubernetes cloud infrastructure, DevOps, CI/CD and improving the developer experience
  • You will be managing cloud deployments in AWS, Azure and on premise
  • This role emphasizes the continuous innovation in improving our cloud computing environments
  • Architect, deploy and maintain infrastructure with cloud providers and Kubernetes (EKS)
  • Collaborate with multi-disciplined teams to define and execute on internal and external deployments
  • Promote SRE best practices in system resilience, performance monitoring and high availability
  • Design, develop, and deliver solutions using infrastructure as code with tools like Terraform and Python
  • Develop and maintain CI/CD pipelines for automated deployment
  • Build strong relationships with internal and external customers to identify technical solutions to their problems
  • Improve Anduril’s operational capabilities by improving our core product offering through root cause analysis and creating tooling capable of managing large scale deployments
  • Lead the organization in building scalable, sustainable mechanisms to continue delivering to customers at the pace the business is scaling
Benefits
  • Full Family Health Coverage
  • 16 Weeks Paid Parental Leave for All Caregivers
  • Family Planning & Support
  • Incentivized Time Off
  • Mental Health Resources
  • Financial Planning
  • Unlimited Provisions
  • Professional Development
Qualifications & Experience

Technical expertise and demonstrated performance in one or more of the following areas: networking, cloud technologies, application development and/or cybersecurity. Experience with cloud services (AWS/Azure). 6+ years of engineering experience. Experience performing data-driven root cause analysis on complex systems. Deep knowledge of the Kubernetes ecosystem (Docker, Helm, ArgoCD, Terraform). Experience in software languages such as Go, Python, Rust, or C++. Eligible to obtain and maintain an active U.S. Secret security clearance. Demonstrated ability to train peers or customers on the operation of a product. Computer Science degree or equivalent. Experience with managing Kubernetes clusters of hundreds of nodes. Knowledge of performance improvement techniques, metrics and alerting. Experience with KubeVirt, qemu, virtualization and hypervisor technologies. Experience with low-level frameworks, Linux and databases. Excellent written and verbal communication skills.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer, TS Clearance
Senior Site Reliability Engineer, TS Clearance

Slope • Washington

On-site
USD 191,000 - 287,000
Comprehensive benefits package
Health and recovery support
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Mosaic.tech • Washington

On-site
USD 166,000 - 220,000
Senior SRE: Cloud & Kubernetes CI/CD for Secure Deployments
Senior SRE: Cloud & Kubernetes CI/CD for Secure Deployments

Mosaic.tech • Washington

On-site
USD 166,000 - 220,000
Senior SRE: Build Scalable Cloud & Kubernetes Platforms
Senior SRE: Build Scalable Cloud & Kubernetes Platforms

Anduril Industries • Washington

On-site
USD 140,000 - 190,000
Full Family Health Coverage
16 Weeks Paid Parental Leave for All-C
Family Planning & Support
+5
Senior Site Reliability Engineer
Senior Site Reliability Engineer

GovCIO • Arlington (VA)

Hybrid
USD 210,000 - 230,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Govcio LLC • United States

Hybrid
USD 210,000 - 230,000
Senior Production Engineer
Senior Production Engineer

Slope • Washington

Hybrid
USD 191,000 - 287,000
Comprehensive benefits package
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Senior Cloud Reliability Engineer — Go, Kubernetes
Senior Cloud Reliability Engineer — Go, Kubernetes

Anduril Industries • United States

On-site
USD 120,000 - 160,000
Full Family Health Coverage
16 Weeks Paid Parental Leave for All Caregivers
Incentivized Time Off
+1
Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • Greenwood Village (CO)

On-site
USD 120,000 - 150,000