SRE, Cloud & Kubernetes Platform Engineer

Algolia

United States

Hybrid

USD 65,000 - 90,000

Full time

9 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Algolia is seeking a Site Reliability Engineer for the IaaS team to help build the next generation of production infrastructure. You will contribute to cloud foundations, Kubernetes, automation, reliability, and large-scale operations across cloud environments.

This is a hands-on role focused on building, operating, and improving scalable systems with emphasis on safety and observability. You will work with cross-functional teams to deliver secure, cost-aware platform capabilities and

Qualifications

  • Hands-on production knowledge of AWS or GCP.
  • Practical Kubernetes knowledge and an interest in operating it in production.
  • Familiarity with infrastructure as code, ideally Terraform.
  • Programming or scripting skills in Python, Go, or an equivalent language.
  • Strong Linux and networking fundamentals.
  • A strong interest in reliability, automation, and solving production problems.
  • Comfort adopting AI-assisted engineering tools, with sound judgement for critical production systems.
  • The ability to communicate clearly and work effectively with a distributed team.
  • Excellent spoken and written English skills.

Responsibilities

  • Build and improve Cloud Baseline capabilities, including identity and access, networking, security, resource inventory, tagging, and auditability.
  • Develop and maintain infrastructure as code and automation for cloud environments and Kubernetes infrastructure.
  • Contribute to reliable, repeatable cloud and cluster lifecycle operations.
  • Help build self-service capabilities, reusable modules, and clear documentation that make the safe path the easy path for platform consumers.
  • Reduce manual work and configuration drift through automation, testing, GitOps practices, and standardisation.
  • Use automation and AI-assisted engineering tools where appropriate to improve infrastructure analysis, documentation, and safe, repeatable changes.
  • Improve observability, monitoring, alerting, capacity management, and operational documentation.
  • Investigate production issues, participate in the on-call rotation, and turn lessons learned into lasting improvements.
  • Work with Infrastructure, Security, FinOps, and engineering teams to deliver reliable, secure, and cost-aware platform capabilities.

Skills

AWS/GCP
Kubernetes
Infrastructure as code
Terraform
Python/Go
Linux
Networking
Reliability
Communication
English

Tools

Argo CD
Helm
OPA
Kyverno

Job description

Algolia is seeking a Site Reliability Engineer for the IaaS team to help build the next generation of production infrastructure. You will contribute to cloud foundations, Kubernetes, automation, reliability, and large-scale operations across cloud environments.

This is a hands-on role focused on building, operating, and improving scalable systems with emphasis on safety and observability. You will work with cross-functional teams to deliver secure, cost-aware platform capabilities and

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE, Cloud Baseline & Kubernetes Platform
Senior SRE, Cloud Baseline & Kubernetes Platform

United States Digital Space LLC • Paris (TX)

On-site
USD 140,000 - 210,000
Senior SRE, AI Platform: Scale Reliability & Kubernetes
Senior SRE, AI Platform: Scale Reliability & Kubernetes

United States Digital Space LLC • Paris (TX)

On-site
USD 80,000 - 111,000
Site Reliability Engineer, Cloud & Kubernetes — Remote
Site Reliability Engineer, Cloud & Kubernetes — Remote

United States Digital Space LLC • Paris (TX)

On-site
USD 140,000 - 190,000
Senior SRE: CI/CD, Kubernetes & Observability Leader
Senior SRE: CI/CD, Kubernetes & Observability Leader

United States Digital Space LLC • Paris (TX)

On-site
USD 80,000 - 111,000
Remote work options
Global offices
Remote SRE for AI Platform - Scale & Reliability
Remote SRE for AI Platform - Scale & Reliability

United States Digital Space LLC • Paris (TX)

On-site
USD 80,000 - 111,000
Platform Reliability Engineer — Kubernetes & Infra Ops
Platform Reliability Engineer — Kubernetes & Infra Ops

Alibaba Cloud • Sunnyvale (CA)

On-site
USD 145,000 - 238,000
Site Reliability Engineer, IaaS
Site Reliability Engineer, IaaS

Algolia • United States

Hybrid
USD 65,000 - 90,000
Global Cloud SRE – Reliability, Security & Observability
Global Cloud SRE – Reliability, Security & Observability

Fideo Intelligence • United States

On-site
USD 90,000 - 100,000
Senior SRE – AI Cloud Platform, Kubernetes Expert
Senior SRE – AI Cloud Platform, Kubernetes Expert

Socket.dev • San Francisco (CA)

On-site
USD 180,000 - 240,000
Health, dental, vision coverage for in
Wellness and commuter stipends
401k with 2% company match
+1
Cloud SRE Specialist - Reliability & Automation Engineer
Cloud SRE Specialist - Reliability & Automation Engineer

Alibaba Cloud • Seattle (WA)

On-site
USD 133,000 - 220,000
Medical, dental, and vision insurance
401(k) plan
Paid holidays and vacation days
+1