Senior DevOps / Site Reliability Engineer

Salt

Abu Dhabi

On-site

AED 480,000 - 720,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Salt is seeking a hands-on Senior DevOps/SRE to build and operate the delivery and runtime foundations for enterprise apps, workflow platforms and AI-enabled products. You will work with software engineers to create reliable, secure, automated paths from code to production, owning deployment automation, observability and operational readiness across integrations with critical systems.

You will design secure environments, manage containerized workloads, define deployment patterns for

Qualifications

  • 6+ years in DevOps, SRE, platform engineering or cloud infrastructure.
  • Docker and Kubernetes.
  • Infrastructure as Code using Terraform, Bicep, Pulumi or equivalent.
  • CI/CD using Azure DevOps, GitHub Actions, GitLab CI or similar.
  • Strong Linux and networking fundamentals.
  • Observability tooling and distributed-system troubleshooting.
  • Secure secrets, identity and access-management patterns.
  • Production incident-management experience.
  • Scripting/programming: Python, Go, Bash or equivalent.

Responsibilities

  • Build and maintain cloud infrastructure using Infrastructure as Code.
  • Design secure, repeatable environments across development, test, staging and production.
  • Manage containerized workloads and Kubernetes-based deployments where appropriate.
  • Define standard application deployment patterns for backend, frontend and AI services.
  • Implement secure secrets and configuration management.
  • Support network, identity and connectivity requirements for enterprise integrations.
  • Build automated CI/CD pipelines and standardize build, test, security scanning and deployment processes.
  • Automate environment provisioning and configuration.
  • Reduce manual deployment steps and production configuration drift.
  • Work with engineering teams to improve release frequency and reliability.
  • Establish logging, metrics, tracing and alerting.
  • Define SLI and operational thresholds.
  • Build dashboards for system health and application performance.
  • Implement incident-response and production-support practices.
  • Design for graceful degradation, retries, failover and recovery.
  • Lead root-cause analysis of production incidents.
  • Implement least-privilege access and secure deployment patterns.
  • Support auditability of infrastructure and production changes.
  • Integrate security checks into delivery pipelines.
  • Meet enterprise control requirements with security and infrastructure teams.
  • Support business continuity and disaster-recovery design.
  • Define backup, restore and recovery procedures.

Skills

DevOps
SRE
Platform engineering
Cloud infrastructure
Linux
Networking
Observability
Incident management
IAM
Python
Go
Bash
CI/CD
Azure
AWS
GCP
Azure

Tools

Docker
Kubernetes
Terraform
Bicep
Pulumi
Azure DevOps
GitHub Actions
GitLab CI

Job description

The Role

We are seeking a hands-on Senior DevOps / Site Reliability Engineer to build and operate the delivery and runtime foundations for a portfolio of modern enterprise applications, workflow platforms and AI-enabled products.

This is not primarily an infrastructure-administration role. You will work directly with software engineers to create reliable, secure and automated paths from code to production.

You will own deployment automation, runtime reliability, observability, infrastructure-as-code and operational readiness across applications that integrate with critical enterprise systems.

What You Will Own
Platform & Infrastructure
  • Build and maintain cloud infrastructure using Infrastructure as Code.
  • Design secure, repeatable environments across development, test, staging and production.
  • Manage containerized workloads and Kubernetes-based deployments where appropriate.
  • Define standard application deployment patterns for backend, frontend and AI services.
  • Implement secure secrets and configuration management.
  • Support network, identity and connectivity requirements for enterprise integrations.
CI/CD & Developer Productivity
  • Build automated CI/CD pipelines.
  • Standardize build, test, security scanning and deployment processes.
  • Automate environment provisioning and configuration.
  • Reduce manual deployment steps and production configuration drift.
  • Work closely with engineering teams to improve release frequency and reliability.
Reliability & Observability
  • Establish logging, metrics, tracing and alerting.
  • Define service-level indicators and operational thresholds.
  • Build dashboards for system health and application performance.
  • Implement incident-response and production-support practices.
  • Design for graceful degradation, retries, failover and recovery.
  • Lead root-cause analysis of production incidents.
Security & Operational Controls
  • Implement least-privilege access and secure deployment patterns.
  • Support auditability of infrastructure and production changes.
  • Integrate security checks into delivery pipelines.
  • Work with security and infrastructure teams to meet enterprise control requirements.
Resilience
  • Support business continuity and disaster-recovery design.
  • Define backup, restore and recovery procedures.
  • Test operational recovery rather than relying solely on documented plans.
Required Experience
  • 6+ years in DevOps, SRE, platform engineering or cloud infrastructure.
  • Strong production experience with Azure, AWS or GCP; Azure strongly preferred.
  • Docker and Kubernetes.
  • Infrastructure as Code using Terraform, Bicep, Pulumi or equivalent.
  • CI/CD using Azure DevOps, GitHub Actions, GitLab CI or similar.
  • Strong Linux and networking fundamentals.
  • Observability tooling and distributed-system troubleshooting.
  • Secure secrets, identity and access-management patterns.
  • Production incident-management experience.
  • Scripting/programming capability in Python, Go, Bash or equivalent.
Strong Advantage
  • Azure Kubernetes Service.
  • Azure Service Bus, API Management, Key Vault and related Azure services.
  • Enterprise integration platforms.
  • SAP-connected environments.
  • AI/LLM application deployment.
  • Regulated or government environments.
  • High-availability and disaster-recovery architecture.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior DevOps / Site Reliability Engineer
Senior DevOps / Site Reliability Engineer

Client of Salt • Abu Dhabi

On-site
AED 320,000 - 520,000
Senior DevOps / Site Reliability Engineer (SRE)
Senior DevOps / Site Reliability Engineer (SRE)

Stellar Technologies • Abu Dhabi

On-site
AED 360,000 - 540,000
Site Reliability Engineer (SRE) - Azure focus
Site Reliability Engineer (SRE) - Azure focus

Dicetek LLC • Dubai

On-site
AED 300,000 - 550,000
Senior Platform Engineer - Cloud & Kubernetes
Senior Platform Engineer - Cloud & Kubernetes

Client of Weekday AI • Abu Dhabi

On-site
AED 350,000 - 600,000
Senior DevOps / SRE Engineer
Senior DevOps / SRE Engineer

GSSTech Group • Dubai

On-site
AED 441,000 - 698,000
Azure Devops Engineer
Azure Devops Engineer

NorthBay Solutions • Abu Dhabi

On-site
AED 293,000 - 441,000
Azure Devops Engineer
Azure Devops Engineer

NorthBay Solutions LLC • Abu Dhabi

On-site
Confidential
Senior Azure DevOps Engineer: Cloud, CI/CD & DevSecOps
Senior Azure DevOps Engineer: Cloud, CI/CD & DevSecOps

NorthBay Solutions • Abu Dhabi

On-site
Confidential
Senior DevSecOps Engineer
Senior DevSecOps Engineer

Epergne Solutions • United Arab Emirates

On-site
AED 201,000 - 312,000
Senior DevOps SRE Engineer
Senior DevOps SRE Engineer

Global Software Solutions Group • Dubai

On-site
AED 320,000 - 520,000