Site Reliability Engineer (SRE) | 8+ Years Experience | Riyadh, Saudi Arabia

Mindpool Technologies Limited

Saudi Arabia

On-site

SAR 240,000 - 360,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Mindpool Technologies Limited in Riyadh, Saudi Arabia is seeking an experienced Site Reliability Engineer (SRE) to join a technical team. The role targets 8+ years in SRE/DevOps and focuses on scalable AWS cloud infrastructure, observability, and automation.

The SRE will manage cloud infrastructure, incident response, RCA, and define SLOs/SLIs while driving reliability through CI/CD, cloud security, and automation best practices.

Qualifications

  • 8+ years of experience in SRE, DevOps, Production Support, or related roles.
  • Strong hands-on experience with AWS Cloud.
  • Practical experience with Kubernetes and Docker.
  • Strong Linux administration and troubleshooting skills.
  • Understanding of networking fundamentals.
  • Experience with monitoring, observability, incident management, and RCA.
  • Proficiency in scripting with Python, Bash, or Go.
  • Experience with production operations and reliability engineering practices.
  • Knowledge of SLOs, SLIs, error budgets, and MTTR concepts.
  • Experience supporting ITIL/ITSM-based environments.
  • Terraform or CloudFormation experience preferred.
  • Jenkins or GitLab CI/CD experience preferred.
  • ServiceNow/ITSM exposure preferred.
  • Knowledge of cloud security tools advantageous.
  • Experience in Cisco or large enterprise environments preferred.
  • Strong analytical, troubleshooting, communication, and problem-solving skills.
  • Available to join within 15 days maximum.

Responsibilities

  • Build and manage scalable AWS cloud infrastructure using Terraform and CloudFormation.
  • Implement and maintain monitoring and observability solutions using Prometheus, Grafana, Splunk, or Datadog.
  • Manage production incidents and provide effective incident resolution.
  • Conduct Root Cause Analysis (RCA) for production issues.
  • Define and manage SLOs, SLIs, and error budgets.
  • Participate in on-call production support.
  • Automate operational activities to improve system reliability and reduce MTTR.
  • Manage Kubernetes and Docker environments.
  • Support and maintain CI/CD pipelines.
  • Administer and troubleshoot Linux environments.
  • Apply networking fundamentals to production infrastructure and troubleshooting.
  • Develop and maintain operational runbooks.
  • Support ITIL and ITSM-based production operations.
  • Collaborate with technical teams to improve system availability, reliability, and operational performance.
  • Support cloud infrastructure upgrades, automation, and reliability initiatives.

Skills

AWS Cloud
Linux Admin
Networking
Scripting (Python/Bash/Go)
Monitoring & Observability
Incident Management
SRE/DevOps Practices
SLOs/SLIs/Error Budgets
CI/CD
ITIL/ITSM familiarity
Cloud Security
Terraform/CloudFormation
Kubernetes
Docker
Jenkins/GitLab CI/CD
ServiceNow/ITSM exposure
Cisco/Enterprise Environments

Tools

Kubernetes
Docker
Terraform
CloudFormation
Jenkins
GitLab CI/CD
ServiceNow/ITSM

Job description

Location

Riyadh, Saudi Arabia

Job Category
  • Information Technology (IT) & Software
  • Engineering & Technical
  • Telecommunications
  • Others / Miscellaneous
Job Overview

We are seeking an experienced Site Reliability Engineer (SRE) to join a technical team in Riyadh, Saudi Arabia. The role is ideal for professionals with 8+ years of experience in SRE, DevOps, production support, or related infrastructure engineering. The successful candidate will work on scalable AWS cloud infrastructure, production reliability, monitoring, automation, and enterprise technology operations.

The SRE will play a key role in improving system reliability and operational efficiency through cloud infrastructure management, observability, incident response, automation, and continuous improvement. The position provides opportunities for career growth, professional development, technical training, and upskilling across AWS, Kubernetes, DevOps, cloud security, and modern reliability engineering practices.

Candidates should have strong hands-on experience with AWS Cloud, Kubernetes, Docker, Linux, networking, scripting, monitoring, incident management, and Root Cause Analysis. The ability to join within a maximum of 15 days is required.

Key Responsibilities
  • Build and manage scalable AWS cloud infrastructure using Terraform and CloudFormation.
  • Implement and maintain monitoring and observability solutions using Prometheus, Grafana, Splunk, or Datadog.
  • Manage production incidents and provide effective incident resolution.
  • Conduct Root Cause Analysis (RCA) for production issues.
  • Define and manage SLOs, SLIs, and error budgets.
  • Participate in on-call production support.
  • Automate operational activities to improve system reliability and reduce MTTR.
  • Manage Kubernetes and Docker environments.
  • Support and maintain CI/CD pipelines.
  • Administer and troubleshoot Linux environments.
  • Apply networking fundamentals to production infrastructure and troubleshooting.
  • Develop and maintain operational runbooks.
  • Support ITIL and ITSM-based production operations.
  • Collaborate with technical teams to improve system availability, reliability, and operational performance.
  • Support cloud infrastructure upgrades, automation, and reliability initiatives.
Requirements & Qualifications
  • 8+ years of experience preferred in SRE, DevOps, Production Support, or related roles.
  • Strong hands-on experience with AWS Cloud.
  • Practical experience with Kubernetes and Docker.
  • Strong Linux administration and troubleshooting skills.
  • Understanding of networking fundamentals.
  • Experience with monitoring, observability, incident management, and Root Cause Analysis.
  • Proficiency in scripting with Python, Bash, or Go.
  • Experience with production operations and reliability engineering practices.
  • Knowledge of SLOs, SLIs, error budgets, and MTTR concepts.
  • Experience supporting ITIL/ITSM-based environments.
  • Terraform or CloudFormation experience is preferred.
  • Jenkins or GitLab CI/CD experience is preferred.
  • ServiceNow/ITSM exposure is preferred.
  • Knowledge of cloud security tools is advantageous.
  • Experience working in Cisco or large enterprise environments is preferred.
  • Strong analytical, troubleshooting, communication, and problem-solving skills.
  • Must be available to join within 15 days maximum.
Salary, Benefits & Career Growth

The job description does not specify a salary or compensation package.

This position offers opportunities for career growth and professional development within Site Reliability Engineering, DevOps, cloud infrastructure, and enterprise production operations. Professionals will gain exposure to AWS, Kubernetes, infrastructure automation, observability, CI/CD, cloud security, incident management, and reliability engineering practices.

The role also provides opportunities to strengthen technical expertise through hands-on experience, continuous upskilling, and industry-relevant cloud and DevOps certifications.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE - AWS Cloud, Kubernetes & Automation
Senior SRE - AWS Cloud, Kubernetes & Automation

Mindpool Technologies Limited • Saudi Arabia

On-site
SAR 240,000 - 360,000
Senior Site Reliability Engineer: Scale, Automation, Observability
Senior Site Reliability Engineer: Scale, Automation, Observability

TAWANTECH • Riyadh

On-site
SAR 240,000 - 320,000
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Jobgether • Saudi Arabia

Remote
SAR 180,000 - 300,000
Fully remote in Saudi Arabia
Competitive compensation
Technical leadership and mentoring
+1
Senior DevOps Engineer
Senior DevOps Engineer

Norconsult Telematics Limited • Saudi Arabia

On-site
SAR 280,000 - 420,000
Senior DevOps & SRE Engineer: Cloud, Automation, Reliability
Senior DevOps & SRE Engineer: Cloud, Automation, Reliability

Norconsult Telematics Limited • Saudi Arabia

On-site
SAR 280,000 - 420,000
Senior Full Stack Developer – React.js, Node.js & Kubernetes | Riyadh, Saudi Arabia
Senior Full Stack Developer – React.js, Node.js & Kubernetes | Riyadh, Saudi Arabia

Esolglobal • Saudi Arabia

On-site
SAR 180,000 - 320,000
Expert Site Reliability Engineer
Expert Site Reliability Engineer

TAWANTECH • Riyadh

On-site
SAR 240,000 - 320,000
Senior Infrastructure Engineer (Saudi National)- Riyadh, KSA
Senior Infrastructure Engineer (Saudi National)- Riyadh, KSA

DS DeepSource • Riyadh

On-site
SAR 200,000 - 320,000
Systems Administrator – DevOps | AAA REZAYAT | Al Khobar, Saudi Arabia
Systems Administrator – DevOps | AAA REZAYAT | Al Khobar, Saudi Arabia

AAA REZAYAT • Al Khobar

On-site
SAR 149,000 - 225,000
Senior Manager - Application Operations
Senior Manager - Application Operations

Rasan • Riyadh

On-site
SAR 300,000 - 480,000