SRE Lead: Cloud Reliability & Observability Architect

Avrioc Technologies

Abu Dhabi

On-site

AED 360,000 - 600,000

Full time

23 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Avrioc Technologies is seeking an experienced Site Reliability Engineer Lead in Abu Dhabi to design, scale, and elevate our cloud infrastructure and observability ecosystem. The role focuses on reliability, performance, and scalability across AWS, GCP, or Azure.

You will architect scalable cloud infrastructure, define SLOs/SLIs, and drive chaos engineering while collaborating with engineering and product teams to embed reliability into development.

Qualifications

  • 8+ years of experience in DevOps/SRE, including leadership in enterprise environments.
  • Hands-on experience with AWS, GCP, or Azure.
  • Strong expertise in Infrastructure as Code (Terraform, CloudFormation, Ansible).
  • Proven experience in CI/CD, monitoring, and incident response.
  • Deep knowledge of observability tools and practices.
  • Strong Kubernetes and Helm experience at scale.
  • Experience with databases like MySQL, Cassandra, etc.
  • Proficiency in Python, Bash, or Go.
  • Experience in BCP/DR planning and capacity management.
  • Strong communication, troubleshooting, and documentation skills.

Responsibilities

  • Architect and deploy scalable, highly available cloud infrastructure
  • Lead SRE best practices to ensure reliability, performance, and scalability
  • Optimize CI/CD pipelines (Jenkins, Argo CD or similar) for seamless deployments
  • Define and track SLOs & SLIs to maintain uptime and service health
  • Build robust observability frameworks (Elastic Stack, Prometheus, Grafana, Dynatrace, New Relic)
  • Manage Kubernetes clusters and Helm charts for efficient orchestration
  • Implement auto-healing systems and proactive monitoring
  • Drive chaos engineering and resilience testing (Chaos Mesh, Litmus, AWS FIS)
  • Collaborate with engineering and product teams to embed reliability into development
  • Maintain clear infrastructure and incident documentation

Skills

SRE leadership
AWS/GCP/Azure
Terraform/CloudFormation/Ansible
CI/CD/Monitoring/Incident response
Kubernetes/Helm
Databases (MySQL/Cassandra)
Python/Bash/Go
BCP/DR planning
Communication & documentation

Tools

Jenkins
Argo CD
Elastic Stack
Prometheus
Grafana
Dynatrace
New Relic

Job description

Avrioc Technologies is seeking an experienced Site Reliability Engineer Lead in Abu Dhabi to design, scale, and elevate our cloud infrastructure and observability ecosystem. The role focuses on reliability, performance, and scalability across AWS, GCP, or Azure.

You will architect scalable cloud infrastructure, define SLOs/SLIs, and drive chaos engineering while collaborating with engineering and product teams to embed reliability into development.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Lead
Site Reliability Engineer - Lead

Avrioc Technologies • Abu Dhabi

On-site
AED 360,000 - 600,000
Senior Site Reliability Engineer — FinTech Cloud & DevOps
Senior Site Reliability Engineer — FinTech Cloud & DevOps

Epergne Solutions • Dubai

On-site
AED 200,000 - 300,000
Senior DevOps & SRE Engineer - Azure, CI/CD, Observability
Senior DevOps & SRE Engineer - Azure, CI/CD, Observability

Stellar Technologies • Abu Dhabi

On-site
AED 360,000 - 540,000
SRE Engineer: Reliability, Observability & Cloud Automation
SRE Engineer: Reliability, Observability & Cloud Automation

Dicetek LLC • Dubai

On-site
AED 300,000 - 520,000
Senior SRE - AIOps & Cloud Reliability Engineer
Senior SRE - AIOps & Cloud Reliability Engineer

DiceTek UAE • Al Ruways Industrial City

On-site
Senior SRE — Space Cloud & SatDevOps Lead
Senior SRE — Space Cloud & SatDevOps Lead

Loft Orbital • Abu Dhabi

On-site
AED 300,000 - 450,000
Senior SRE — Space Ground Systems & Cloud Infra
Senior SRE — Space Ground Systems & Cloud Infra

Loft Orbital Solutions • Abu Dhabi

On-site
AED 450,000 - 750,000
Head of SRE: Cloud Reliability & Platform Engineering
Head of SRE: Cloud Reliability & Platform Engineering

Client of Mark Williams • Dubai

On-site
AED 600,000 - 1,200,000
Azure DevOps SRE Lead – Fintech Reliability & Cloud
Azure DevOps SRE Lead – Fintech Reliability & Cloud

Happiest Minds Technologies • Abu Dhabi

On-site
AED 260,000 - 520,000
Senior SRE: Cloud, Kubernetes & Automation Lead
Senior SRE: Cloud, Kubernetes & Automation Lead

Good co India • United Arab Emirates

On-site
AED 240,000 - 480,000