Turn this role into an interview — a resume and cover letter built around what this employer wants.
Innovations Global is urgently seeking an experienced Site Reliability Engineer (SRE) to champion availability, scalability, and performance for banking platforms in Abu Dhabi on a 1-year renewable contract. You will bridge software engineering and operations, lead observability, automation, incident recovery, and enforce SRE principles across hybrid cloud environments.
The role requires 5+ years in SRE/DevOps, strong Linux/Unix, Docker, Kubernetes, and cloud experience.
Innovations Global is urgently seeking an experienced, proactive Site Reliability Engineer (SRE) to champion the continuous availability, scalability, and performance of mission-critical banking and financial technology platforms in Abu Dhabi, United Arab Emirates. Operating in a full-time onsite capacity for a 1-year renewable contract, you will bridge the divide between software engineering and systems operations to ensure enterprise infrastructure resilience. With a minimum of five years of hands-on experience, you will spearhead observability, drive end-to-end automation, lead rapid incident recovery, and enforce stringent Site Reliability Engineering principles across hybrid cloud environments. This position provides an exceptional opportunity for a technical engineering professional to make a transformative impact on premier banking infrastructure in the UAE.
As a Site Reliability Engineer (SRE) at Innovations Global supporting a premier enterprise banking environment in Abu Dhabi, you will hold operational accountability for the health, availability, performance, and efficiency of high-throughput transactional applications and distributed systems. You will collaborate closely with cross-functional software engineering teams, DevOps squads, database administrators, and cyber security teams to establish resilient deployment pipelines and maintain robust production ecosystems.
Your core technical mandate involves architecting and administering enterprise Linux and Unix servers, orchestrating microservices utilizing Docker and Kubernetes, and managing scalable workloads across leading cloud platforms (AWS, Azure, or GCP). You will implement proactive monitoring and observability frameworks, define and track Service Level Objectives (SLOs), Service Level Indicators (SLIs), and Service Level Agreements (SLAs), and eliminate operational toil through Python and Bash automation scripting. Additionally, you will direct incident response triage, execute root cause analysis (RCA), optimize CI/CD release workflows, and troubleshoot complex TCP/IP enterprise networking bottlenecks. This role requires rigorous diagnostic discipline, deep systems acumen, and the capability to maintain zero-downtime reliability within a fast-paced financial services environment.
Given that this role supports an enterprise banking environment in Abu Dhabi with an immediate-to-30-day onboarding window, highlight your direct financial services or high-concurrency transactional systems background on page one of your CV. Clearly articulate the specific observability stacks (Prometheus, Grafana, ELK) you have architected, quantify the repetitive toil you eliminated via Python or Bash scripting, and explicitly note your current location and notice period when emailing sp.hariharan@innovationsglobal.com.