Senior SRE & DevOps - Agent-First Reliability Engineer

Dover

Northern (KY)

Hybrid

USD 120,000 - 160,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Kintsugi is seeking a Senior Site Reliability Engineer (DevOps) to scale and harden production infrastructure across managed Kubernetes and AWS. You’ll own reliability, reduce toil, and build tooling that lets engineers move fast without sacrificing stability.

You will collaborate with Platform Engineering, Product, and QA to design resilient architectures, improve deployment pipelines, and extend internal tools that support engineering velocity while preserving security and observability.

Qualifications

  • 5-8 years in Site Reliability Engineering, DevOps, or Infrastructure Engineering roles with ownership of a production system at meaningful scale
  • Fluent agent-first daily work — directing coding agents to do real engineering work
  • Strong foundation in AWS-hosted data & networking (RDS/Postgres, ElastiCache/Redis, VPC) and managed Kubernetes
  • Track record of building tools that remove manual work (scripts, services, internal platforms)
  • Hands-on experience with CI/CD pipelines and infrastructure-as-code (Terraform, CloudFormation)
  • Observability expertise (metrics, tracing, logging) and modern monitoring practices
  • Familiarity with cloud security/compliance (SOC 2, GDPR a plus)
  • Collaborative mindset to empower developers to move fast safely
  • Experience with developer enablement — internal tooling and platforms for productivity
  • Nice to have: multi-cloud operations across providers

Responsibilities

  • Own the reliability of production infrastructure on Kubernetes and AWS
  • Work agent-first day to day: build, debug, and automate with agentic coding workflows
  • Build internal tools and automation to eliminate recurring toil
  • Develop and operate monitoring, alerting, and observability across the stack
  • Partner with engineering teams to design for reliability and performance from the start
  • Automate infrastructure management via infrastructure-as-code and improve CI/CD
  • Lead and evolve incident response, including postmortems and blameless learning
  • Optimize infrastructure for cost efficiency while maintaining high availability and security
  • Contribute to security, compliance, and disaster recovery efforts
  • Support developer enablement with in-house tooling and internal platforms

Skills

Agent-first mindset
AWS & Kubernetes
CI/CD pipelines
Observability
Infrastructure as code
Security/compliance awareness

Tools

Terraform
CloudFormation
Kubernetes

Job description

Kintsugi is seeking a Senior Site Reliability Engineer (DevOps) to scale and harden production infrastructure across managed Kubernetes and AWS. You’ll own reliability, reduce toil, and build tooling that lets engineers move fast without sacrificing stability.

You will collaborate with Platform Engineering, Product, and QA to design resilient architectures, improve deployment pipelines, and extend internal tools that support engineering velocity while preserving security and observability.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (DevOps)
Senior Site Reliability Engineer (DevOps)

Dover • Northern (KY)

Hybrid
USD 120,000 - 160,000
Senior SRE: Build Resilient AWS/Kubernetes Infrastructure
Senior SRE: Build Resilient AWS/Kubernetes Infrastructure

Kontakt.io • United States

On-site
USD 150,000 - 190,000
Senior SRE – Cloud-native, Kubernetes & CI/CD
Senior SRE – Cloud-native, Kubernetes & CI/CD

Pinterest • San Francisco (CA)

On-site
USD 139,764 - 287,749
Equity
Competitive salary
Senior SRE: Kubernetes, Cloud Reliability & Automation
Senior SRE: Kubernetes, Cloud Reliability & Automation

DraftKings • Boston (MA)

On-site
USD 128,000 - 160,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kovoro • Denver (CO), Northern (KY)

On-site
USD 150,000 - 190,000
Senior SRE - Multi-Region AWS, Kubernetes & CI/CD
Senior SRE - Multi-Region AWS, Kubernetes & CI/CD

Grubhub Holdings Inc. • New York (NY)

On-site
USD 209,000 - 217,000
Equity
401(k)
Medical/Dental/Vision
Senior SRE with TS Clearance – Resilience & Automation
Senior SRE with TS Clearance – Resilience & Automation

Peerless Technologies Corporation • Dayton (OH)

On-site
USD 120,000 - 180,000
Senior Site Reliability Engineer, Forward Deployed - Remote USA ONLY
Senior Site Reliability Engineer, Forward Deployed - Remote USA ONLY

Ardan Labs • United States

On-site
USD 140,000 - 210,000
Senior SRE: Cloud, Kubernetes & Automation
Senior SRE: Cloud, Kubernetes & Automation

Socure • Carson City (NV)

On-site
USD 150,000 - 190,000
Staff SRE: Scale, Observability & Kubernetes
Staff SRE: Scale, Observability & Kubernetes

Replit • Foster City (CA)

On-site
USD 180,000 - 260,000
Competitive Salary & Equity
401(k) with 4% match
Health, Dental, Vision and Life Ins.
+2