Site Reliability Engineer (SRE)

Dicetek LLC

Dubai

On-site

AED 480,000 - 780,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Dicetek LLC in Dubai seeks an experienced Site Reliability Engineer to ensure reliability, availability, and performance of critical applications and infrastructure. You will design and maintain monitoring, CI/CD pipelines, and IaC, while partnering with development and operations teams.

The ideal candidate has hands-on experience with cloud platforms (AWS/Azure/GCP), containers, and scripting, plus strong incident management and collaboration skills.

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, or a related field.
  • Proven experience as a Site Reliability Engineer, DevOps Engineer, or Production Support Engineer.
  • Strong experience in cloud infrastructure, automation, monitoring, and incident management.
  • Hands-on experience with Kubernetes and containerized environments.

Responsibilities

  • Monitor and maintain the availability, performance, and reliability of applications and infrastructure.
  • Design and implement effective monitoring, logging, and alerting solutions.
  • Manage and troubleshoot production incidents and perform root cause analysis.
  • Automate operational and deployment processes to improve system reliability and efficiency.
  • Work closely with development, infrastructure, and DevOps teams to improve application performance and resilience.
  • Implement and maintain CI/CD pipelines and Infrastructure as Code (IaC).
  • Support containerized environments and cloud-based infrastructure.
  • Develop scripts and automation tools to reduce manual operational activities.
  • Implement observability solutions, including monitoring, logging, and distributed tracing.
  • Ensure proper documentation of operational procedures, incidents, and system configurations.

Skills

Monitoring & observability
Incident management
DevOps practices
Collaboration

Education

Bachelor's degree in CS/IT or related field

Tools

Prometheus
Grafana
Zabbix
ELK/Elastic Stack
Splunk
Dynatrace
AppDynamics
New Relic
AWS
Azure
GCP
Docker
Kubernetes
OpenShift
Jenkins
GitLab CI/CD
Azure DevOps
Terraform
Ansible
Git
GitHub
GitLab
ServiceNow
PagerDuty
OpenTelemetry
Jaeger
Bash
Python
PowerShell
PostgreSQL
Oracle
SQL Server
IIS
Nginx
Apache
REST APIs

Job description

Job Purpose

We are seeking an experienced Site Reliability Engineer (SRE) to ensure the reliability, availability, performance, and scalability of critical applications and infrastructure. The ideal candidate will have strong experience in monitoring, automation, cloud technologies, incident management, and DevOps practices.

Key Responsibilities
  • Monitor and maintain the availability, performance, and reliability of applications and infrastructure.
  • Design and implement effective monitoring, logging, and alerting solutions.
  • Manage and troubleshoot production incidents and perform root cause analysis.
  • Automate operational and deployment processes to improve system reliability and efficiency.
  • Work closely with development, infrastructure, and DevOps teams to improve application performance and resilience.
  • Implement and maintain CI/CD pipelines and Infrastructure as Code (IaC).
  • Support containerized environments and cloud-based infrastructure.
  • Develop scripts and automation tools to reduce manual operational activities.
  • Implement observability solutions, including monitoring, logging, and distributed tracing.
  • Ensure proper documentation of operational procedures, incidents, and system configurations.
Required Technical Skills
  • Monitoring: Prometheus, Grafana, Zabbix
  • Logging: ELK/Elastic Stack, Splunk
  • APM: Dynatrace, AppDynamics, New Relic
  • Cloud: AWS, Microsoft Azure, or GCP
  • Containers: Docker, Kubernetes, OpenShift
  • CI/CD: Jenkins, GitLab CI/CD, Azure DevOps
  • Infrastructure as Code: Terraform, Ansible
  • Version Control: Git, GitHub, GitLab
  • Incident Management: ServiceNow, PagerDuty
  • Distributed Tracing: OpenTelemetry, Jaeger
  • Scripting: Bash, Python, PowerShell
  • Databases: PostgreSQL, Oracle, SQL Server
  • Web & APIs: IIS, Nginx, Apache, REST APIs
Qualifications & Experience
  • Bachelor's degree in Computer Science, Information Technology, or a related field.
  • Proven experience as a Site Reliability Engineer, DevOps Engineer, or Production Support Engineer.
  • Strong experience in cloud infrastructure, automation, monitoring, and incident management.
  • Hands-on experience with Kubernetes and containerized environments.
  • Strong troubleshooting and root cause analysis skills.
  • Experience working in highly available, large-scale production environments.
  • Excellent communication and collaboration skills.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE) - Azure focus
Site Reliability Engineer (SRE) - Azure focus

Dicetek LLC • Dubai

On-site
AED 300,000 - 550,000
SRE (Site Reliability Engineer)
SRE (Site Reliability Engineer)

Dicetek LLC • Abu Dhabi

On-site
AED 180,000 - 300,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Epergne Solutions • Abu Dhabi

On-site
AED 180,000 - 250,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

31 CONCEPT • United Arab Emirates

On-site
AED 300,000 - 460,000
Senior DevOps / SRE Engineer
Senior DevOps / SRE Engineer

Legend Holding Group Ltd • Dubai

On-site
AED 450,000 - 650,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Dicetek LLC • Abu Dhabi

On-site
AED 180,000 - 280,000
Senior DevOps / Site Reliability Engineer
Senior DevOps / Site Reliability Engineer

Salt • Abu Dhabi

On-site
AED 480,000 - 720,000
Senior DevOps / Site Reliability Engineer
Senior DevOps / Site Reliability Engineer

Client of Salt • Abu Dhabi

On-site
AED 320,000 - 520,000
Lead Site Reliability Engineer at HCLTech
Lead Site Reliability Engineer at HCLTech

HCLTech • United Arab Emirates

On-site
AED 120,000 - 160,000
Senior DevOps / SRE Engineer
Senior DevOps / SRE Engineer

GSSTech Group • Dubai

On-site
AED 441,000 - 698,000