Site Reliability Engineer: Scalable Systems & Automation

Dicetek LLC

Abu Dhabi

On-site

AED 180,000 - 300,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Dicetek LLC is seeking an experienced Site Reliability Engineer to design, build, and operate scalable, highly available production systems in our Abu Dhabi environment. You will automate operational tasks, implement robust monitoring and logging, and participate in incident response and post-mortems to prevent recurrence, ensuring SLA/SLO/SLI alignment and continuous improvement across platforms.

The ideal candidate has strong Linux/Unix administration, cloud platform experience

Qualifications

  • Strong Linux/Unix administration and performance tuning.
  • Hands-on cloud experience with AWS, Azure or GCP.
  • Proficient in Kubernetes, Docker, automation and CI/CD pipelines.
  • Experience with monitoring, observability, incident management and RCA.
  • Banking/fintech or enterprise app exposure is a plus.

Responsibilities

  • Design, build, and maintain scalable and highly available production systems.
  • Automate operational tasks via scripting and tool development.
  • Implement robust monitoring, alerting, and logging solutions.
  • Participate in incident response and post-mortems to identify root causes.

Skills

Linux
cloud platforms
Kubernetes
Docker
monitoring & observability
scripting (Python/Bash)
CI/CD
networking
incident management
automation
troubleshooting
SRE practices (SLA/SLO/SLI)

Tools

Docker
Kubernetes
CI/CD pipelines tooling

Job description

Dicetek LLC is seeking an experienced Site Reliability Engineer to design, build, and operate scalable, highly available production systems in our Abu Dhabi environment. You will automate operational tasks, implement robust monitoring and logging, and participate in incident response and post-mortems to prevent recurrence, ensuring SLA/SLO/SLI alignment and continuous improvement across platforms.

The ideal candidate has strong Linux/Unix administration, cloud platform experience

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Automation & Observability
Site Reliability Engineer - Automation & Observability

Dicetek LLC • Dubai

On-site
AED 300,000 - 550,000
Senior SRE - AIOps & Cloud Reliability Engineer
Senior SRE - AIOps & Cloud Reliability Engineer

DiceTek UAE • Al Ruways Industrial City

On-site
AI Platform SRE: Resilience & Automation
AI Platform SRE: Resilience & Automation

Dicetek LLC • Abu Dhabi

Hybrid
AED 120,000 - 180,000
Site Reliability Engineer - AIOps
Site Reliability Engineer - AIOps

DiceTek UAE • Al Ruways Industrial City

On-site
Senior Infrastructure Architect - Cloud & HA
Senior Infrastructure Architect - Cloud & HA

Dicetek LLC • Dubai

On-site
AED 335,000 - 614,000
System Support Engineer - Dicetek LLC - Dubai, UAE
System Support Engineer - Dicetek LLC - Dubai, UAE

DiceTek UAE • Dubai

On-site
Opportunity to upgrade skills in cloud and virtualization
Stable, long-term role with strong career development potential
Supportive technical environment with continuous learning
System Engineer - Dicetek LLC - Dubai, UAE
System Engineer - Dicetek LLC - Dubai, UAE

DiceTek UAE • Dubai

On-site
Work with enterprise-grade infrastructure technologies
Gain exposure to complex IT environments
Supportive, growth-focused technical team
+1
Lead DevOps Engineer - Cloud, CI/CD & Kubernetes
Lead DevOps Engineer - Cloud, CI/CD & Kubernetes

Dicetek LLC • Dubai

On-site
AED 360,000 - 540,000
Senior Linux Systems Engineer: Automation & Reliability
Senior Linux Systems Engineer: Automation & Reliability

DeepSource Technologies • United Arab Emirates

On-site
AED 120,000 - 240,000
Cloud Operations Engineer: Drive Reliability & Automation
Cloud Operations Engineer: Drive Reliability & Automation

DiceTek UAE • Dubai

On-site
Exposure to cloud operations and incident management
Opportunity to grow in cloud support roles
Work in stable and reliable environments