Senior SRE: Scalable Kubernetes Platform & IaC Architect

Circle

San Francisco (CA)

On-site

USD 123,000 - 205,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Circle is seeking a Senior Site Reliability Engineer to design, build, and operate the scalable platform infrastructure behind critical digital assets, AI, and app workloads. You will write and maintain services, automate repeatable workflows, and advance Kubernetes-based platforms across hybrid and public-cloud environments.

You will collaborate with platform, product, and application teams to translate workload requirements into resilient designs, taking ownership of production outcomes and

Qualifications

  • 5+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a closely related software engineering role supporting production systems.
  • Deep, hands-on Kubernetes expertise: designing, operating, securing, and troubleshooting production clusters and containerized workloads at scale.
  • Strong Terraform experience, including authoring reusable modules, managing state and environments, and delivering infrastructure changes through reviewable, automated workflows.

Responsibilities

  • Design, build, and operate Kubernetes platforms that provide secure, highly available, and scalable foundations for production services across hybrid and public-cloud environments.
  • Build infrastructure as code with Terraform, creating reusable modules, safe delivery workflows, and well-governed infrastructure changes.
  • Develop backend services, internal tools, and operational automation in Go, Python, or JavaScript/TypeScript to eliminate manual work and improve the developer experience.
  • Partner with engineering and product teams to understand workload requirements and design pragmatic solutions for reliability, performance, capacity, security, and cost.
  • Improve the production lifecycle through reliable CI/CD, deployment automation, progressive delivery, and clear operational ownership.
  • Define and evolve observability practices across metrics, logs, traces, alerting, and dashboards so teams can detect issues early and troubleshoot effectively.
  • Own production reliability by participating in on-call, incident response, root-cause analysis, and durable corrective actions.
  • Establish and maintain reliability targets through SLIs, SLOs, error budgets, capacity planning, disaster-recovery testing, and resilience improvements.
  • Embed security and compliance into platform operations, partnering with Security to protect infrastructure, workloads, and data while meeting regulatory requirements.
  • Apply AI-assisted and data-driven operational techniques to improve signal detection and automation opportunities.
  • Raise the bar through code reviews, documentation, knowledge sharing, and mentorship.
  • Mentor and support team growth, fostering collaboration and scalability.

Skills

Ownership
Strong communication
Problem solving

Tools

Kubernetes
Terraform
Go
Python
JavaScript/TypeScript
CI/CD
GitOps
Cloud platforms
Observability

Job description

Circle is seeking a Senior Site Reliability Engineer to design, build, and operate the scalable platform infrastructure behind critical digital assets, AI, and app workloads. You will write and maintain services, automate repeatable workflows, and advance Kubernetes-based platforms across hybrid and public-cloud environments.

You will collaborate with platform, product, and application teams to translate workload requirements into resilient designs, taking ownership of production outcomes and

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer - Infra Ops
Senior Site Reliability Engineer - Infra Ops

Circle • San Francisco (CA)

On-site
USD 123,000 - 205,000
Senior Site Reliability Engineer: AI-Powered Cloud Platform
Senior Site Reliability Engineer: AI-Powered Cloud Platform

Circle • San Francisco (CA)

On-site
USD 153,000 - 205,000
Senior SRE, DevEx Platform & Kubernetes Automation
Senior SRE, DevEx Platform & Kubernetes Automation

Chainlink Labs • Las Vegas (NV)

On-site
USD 150,000 - 180,000
Senior SRE: Cloud, Kubernetes & Automation Lead
Senior SRE: Cloud, Kubernetes & Automation Lead

Clover Health • United States

On-site
USD 140,000 - 190,000
401(k) matching
Medical, dental, vision coverage
No-Meeting Fridays
+2
Senior SRE Lead: Platform Reliability & Kubernetes
Senior SRE Lead: Platform Reliability & Kubernetes

Hobbsnews • Chandler (AZ), Northern (KY)

Hybrid
USD 120,000 - 180,000
Senior SRE: Incident & Kubernetes Reliability Lead
Senior SRE: Incident & Kubernetes Reliability Lead

Unique System Skills • United States

Remote
USD 140,000 - 190,000
Senior SRE II — Scale Systems with AI-Driven Reliability
Senior SRE II — Scale Systems with AI-Driven Reliability

Juniper Square • United States

On-site
USD 165,000 - 195,000
Health, dental, and vision care
Life insurance
Mental wellness coverage
+3
Senior SRE – AI Cloud Platform, Kubernetes Expert
Senior SRE – AI Cloud Platform, Kubernetes Expert

Socket.dev • San Francisco (CA)

On-site
USD 180,000 - 240,000
Health, dental, vision coverage for in
Wellness and commuter stipends
401k with 2% company match
+1
Senior SRE - Cloud Infra, AWS & Kubernetes
Senior SRE - Cloud Infra, AWS & Kubernetes

Kids for the Future • United States

Hybrid
USD 130,000 - 195,000
Healthcare benefits
401(k) matching
Paid time off
Senior SRE: Kubernetes, Cloud Reliability & Automation
Senior SRE: Kubernetes, Cloud Reliability & Automation

DraftKings • Boston (MA)

On-site
USD 128,000 - 160,000