Senior Cloud Reliability Engineer: AWS & Kubernetes

Salve.Inno Consulting

United States

On-site

USD 140,000 - 190,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Fully remote position
Long-term cooperation
Work with enterprise customers
High level of technical ownership

Job summary

Salve.Inno Consulting is seeking a Senior Cloud Reliability Engineer to own and operate highly available cloud-native environments. You will work across AWS and Kubernetes with an emphasis on SRE practices, observability, automation, and on-call incident management.

Role offers remote work with international collaboration on complex, production-critical systems. You will engage directly with enterprise customers and influence cloud architecture, reliability practices, and platform improvements

Qualifications

  • Extensive hands-on SRE experience in cloud environments and production systems.
  • Senior AWS expertise in designing, troubleshooting, and improving architectures.
  • Proven Kubernetes production experience, ideally with Amazon EKS.
  • Strong IaC background with Terraform and configuration automation.
  • Experience on-call rotations and ownership of major incidents.
  • Excellent communication with customers and engineering teams.

Responsibilities

  • Own reliability and operational health of production AWS and Kubernetes environments.
  • Participate in on-call rotations and manage high-severity incidents from detection to resolution.
  • Lead root cause analyses and post-incident reviews with permanent remediation.
  • Identify architectural and security weaknesses and implement improvements.
  • Propose and implement improvements based on AWS and Kubernetes best practices.
  • Design, maintain, and improve AWS infra and production Kubernetes/EKS environments.
  • Automate provisioning and workflows using Terraform and config-management tools.
  • Improve deployment and GitOps processes with Argo CD and related tools.
  • Develop observability with Prometheus, Grafana, ELK; build dashboards and alerts.
  • Scale workloads using Kubernetes autoscaling (Karpenter/KEDA).
  • Support distributed, event-driven environments including Kafka.
  • Develop automation tooling with Python, Bash, Go or similar.
  • Strengthen cloud security, DR, and production readiness practices.
  • Collaborate with enterprise customers and engineering teams on troubleshooting and improvements.

Skills

Site Reliability Engineering
Cloud reliability
Platform engineering
AWS
Kubernetes / EKS
Terraform
Ansible
CI/CD
GitOps
Python
Bash
Go
Incident management
On-call experience
Observability
Prometheus
Grafana
ELK
Kafka
Communication in English

Education

AWS certification

Tools

Argo CD
Karpenter
KEDA
Prometheus
Grafana
ELK
Kafka
Terraform
Ansible

Job description

Salve.Inno Consulting is seeking a Senior Cloud Reliability Engineer to own and operate highly available cloud-native environments. You will work across AWS and Kubernetes with an emphasis on SRE practices, observability, automation, and on-call incident management.

Role offers remote work with international collaboration on complex, production-critical systems. You will engage directly with enterprise customers and influence cloud architecture, reliability practices, and platform improvements

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Reliability Engineer (AWS & Kubernetes)
Senior Cloud Reliability Engineer (AWS & Kubernetes)

Salve.Inno Consulting • Chicago (IL)

On-site
USD 130,000 - 190,000
Senior Cloud & Reliability Engineer
Senior Cloud & Reliability Engineer

Very Good Security • United States

Hybrid
USD 100,000 - 130,000
Flexible PTO
Health benefits
401k with matching
+2
Senior Cloud SRE: Automation, Observability & Reliability
Senior Cloud SRE: Automation, Observability & Reliability

Salve.Lab • Atlanta (GA)

On-site
USD 140,000 - 190,000
Remote-friendly
Professional development
Apple equipment
+1
Remote Senior Cloud SRE - AWS & Kubernetes, CI/CD
Remote Senior Cloud SRE - AWS & Kubernetes, CI/CD

SitusAMC • Frankfort (KY)

Remote
USD 95,000 - 135,000
PTO & paid holidays
401K plan
Medical, dental, vision, life and/or 0
Senior Cloud SRE - AWS, Kubernetes & CI/CD (Remote)
Senior Cloud SRE - AWS, Kubernetes & CI/CD (Remote)

SitusAMC • Montpelier (VT)

Remote
USD 95,000 - 135,000
PTO & Holidays
Medical, Dental, Vision
Remote Site Reliability Engineer: Cloud, Kubernetes & Incidents
Remote Site Reliability Engineer: Cloud, Kubernetes & Incidents

AssureSoft Corporation • United States

Remote
USD 120,000 - 180,000
Great Place To Work certification
English scholarships
English classes with company teachers
+3
Remote Senior Site Reliability Lead - AWS, Kubernetes, CI/CD
Remote Senior Site Reliability Lead - AWS, Kubernetes, CI/CD

Empower • United States

On-site
USD 114,000 - 166,000
401(k) with company matching
Tuition reimbursement
Paid volunteer time
+1
Senior Cloud Reliability Engineer — Remote/Hybrid
Senior Cloud Reliability Engineer — Remote/Hybrid

Loftware • United States

Hybrid
USD 115,000 - 160,000
Senior Site Reliability Engineer – Cloud, Kubernetes & Automation
Senior Site Reliability Engineer – Cloud, Kubernetes & Automation

Socure • United States

On-site
USD 160,000 - 180,000
Senior Cloud SRE & Platform Reliability Engineer
Senior Cloud SRE & Platform Reliability Engineer

Convergys • Seattle (WA)

On-site
USD 115,000 - 140,000
Medical insurance
Dental insurance
Vision insurance
+2