Kubernetes SRE Lead: Scalable Cloud Platform on AWS

Okta

San Francisco (CA)

Hybrid

USD 194,000 - 267,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Health, dental, and vision insurance
401(k)
Flexible spending account
Paid leave including PTO and parental leave

Job summary

A leading technology company is seeking a Site Reliability Engineer to build and manage scalable Kubernetes platforms on AWS. The ideal candidate will ensure high availability and performance while optimizing costs. Key qualifications include extensive experience with Kubernetes, AWS, and Terraform, alongside a strong foundation in automation and CI/CD processes. This role is integral to maintaining the reliability and efficiency of cloud-native applications, offering competitive compensation and a vibrant company culture.

Qualifications

  • 4+ years of experience with Kubernetes/Helm.
  • 4+ years of experience with Terraform.
  • 5+ years of experience with AWS.
  • Experience with multi-region cloud environments.
  • Strong expertise in Kubernetes platform creation, management, and optimisation.

Responsibilities

  • Design, implement, and maintain Kubernetes platforms.
  • Build, manage, and optimize AWS cloud infrastructure.
  • Utilize Helm to automate application deployments.
  • Implement and manage Karpenter for scaling.
  • Automate deployment, scaling, and management of infrastructure.

Skills

Kubernetes/Helm
Terraform
AWS
CI/CD pipelines
Scripting in Python, Bash, or Go

Education

Bachelor's degree in Computer Science, Engineering, or related field

Tools

Prometheus
Grafana
CloudWatch
ELK Stack

Job description

A leading technology company is seeking a Site Reliability Engineer to build and manage scalable Kubernetes platforms on AWS. The ideal candidate will ensure high availability and performance while optimizing costs. Key qualifications include extensive experience with Kubernetes, AWS, and Terraform, alongside a strong foundation in automation and CI/CD processes. This role is integral to maintaining the reliability and efficiency of cloud-native applications, offering competitive compensation and a vibrant company culture.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer: Cloud, Kubernetes & CI/CD
Senior Site Reliability Engineer: Cloud, Kubernetes & CI/CD

Amiri Recruiting • Mountain View (CA)

On-site
USD 130,000 - 160,000
Kubernetes SRE — Cloud Platform Reliability & Automation
Kubernetes SRE — Cloud Platform Reliability & Automation

Okta • Chicago (IL)

On-site
USD 174,000 - 214,000
Equity
Health insurance
Paid leave
+2
Site Reliability Engineer - Kubernetes & Cloud
Site Reliability Engineer - Kubernetes & Cloud

Hydrolix • United States

On-site
USD 110,000 - 150,000
Cloud-Native SRE Lead: Automation, Kubernetes & Resilience
Cloud-Native SRE Lead: Automation, Kubernetes & Resilience

Axiom Pursuits • San Francisco (CA)

On-site
USD 150,000 - 180,000
Senior Site Reliability Engineer: Cloud, Kubernetes Uptime
Senior Site Reliability Engineer: Cloud, Kubernetes Uptime

Compunnel, Inc. • Greenwood Village (CO)

On-site
USD 120,000 - 150,000
SRE: AI/ML Infra on Kubernetes, AWS & Terraform
SRE: AI/ML Infra on Kubernetes, AWS & Terraform

Deepgram • United States

Hybrid
USD 120,000 - 150,000
Medical, dental, vision benefits
Unlimited PTO
Generous paid parental leave
+1
Senior Java SRE & Platform Engineer - AWS/Kubernetes
Senior Java SRE & Platform Engineer - AWS/Kubernetes

EITACIES Inc. • Santa Clara (CA)

On-site
USD 120,000 - 160,000
SRE Engineer: Cloud, Kubernetes & CI/CD Reliability
SRE Engineer: Cloud, Kubernetes & CI/CD Reliability

OPPO • Palo Alto (CA)

On-site
USD 100,000 - 200,000
SRE/DevOps Platform Engineer (AWS, Kubernetes)
SRE/DevOps Platform Engineer (AWS, Kubernetes)

Engtal • Washington

On-site
USD 100,000 - 130,000
Equity in the form of stock options
Flexible time-off policy and company holidays
Health, dental, and vision insurance with company contributions
+3
Senior SRE: Cloud-Native Reliability & Automation
Senior SRE: Cloud-Native Reliability & Automation

Crypto Pro Network • New York (NY)

On-site
USD 150,000 - 180,000