Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Dicetek LLC is seeking an experienced Site Reliability Engineer to design, build, and operate scalable, highly available production systems in our Abu Dhabi environment. You will automate operational tasks, implement robust monitoring and logging, and participate in incident response and post-mortems to prevent recurrence, ensuring SLA/SLO/SLI alignment and continuous improvement across platforms.
The ideal candidate has strong Linux/Unix administration, cloud platform experience
SRE (Site Reliability Engineer):
Responsible for ensuring application and infrastructure reliability, availability, performance, and operational efficiency through monitoring, automation, and incident management.
Linux administration, cloud platforms (AWS/Azure/GCP), Kubernetes & Docker, monitoring and observability, scripting (Python/Bash), CI/CD, networking, incident management, automation, troubleshooting, and reliability engineering practices (SLA/SLO/SLI).
Linux
cloud platforms
Kubernetes