Principal SRE

Professional.me

Abu Dhabi

On-site

AED 400,000 - 600,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Alpheya in Abu Dhabi seeks a senior SaaS Operations leader to own end-to-end service operations, from Kubernetes clusters to executive reviews. You will define SLAs, drive incident management, and scale a multi-tenant platform with ISO27001/SOC2 aligned audits.

You will report to the CTO, lead technology-driven reliability, and coordinate regional DR, data residency, and cost modeling across tenants.

Qualifications

  • 12+ years in production operations, including multi-tenant SaaS.
  • Ownership of customer-facing service management and post-incident reviews.
  • Led an SRE/platform-ops team and designed on-call across regions.
  • Deep Kubernetes and cloud knowledge to challenge DR design and capacity claims.
  • Working knowledge of Azure regions, networking, identity, and quota planning.
  • Experience with ISO 27001, SOC 2, or bank outsourcing regulation audits.

Responsibilities

  • Service management for every SaaS customer: SLAs, incident communications, post-incident reviews, audits.
  • Release and deployment operations across tenants: promote environments and change management.
  • Reliability program: DR tests, backups, capacity planning, cost-per-tenant model.
  • Tenant onboarding runbook to go-live on a repeatable path.
  • Europe launch with new Azure regions, DR, residency, and support coverage.
  • Compliance operations with risk teams under ISO 27001, SOC 2, and banking regulations.

Skills

SRE leadership
Incident management
On-call rotation management
Kubernetes depth
Azure cloud
Audit & compliance coordination

Tools

Kubernetes
Azure
Cloud platforms

Job description

Location: Abu Dhabi, UAE

About the Client

Alpheya is a wealth-management technology company headquartered in Abu Dhabi. Banks are their customers. They build each one a tailored investing experience, from mobile apps to advisor and back-office portals, on a single shared platform covering the full order-to-custody lifecycle. The platform runs as SaaS on Microsoft Azure, with on-premises delivery for banks that require it.

About the Role

Over the next six months we are taking multiple banks live as SaaS customers. Our engineering teams build the platform and our SREs run the infrastructure. What we don't yet have is a single leader accountable for the service our customers buy: the SLAs, the incident review a bank CIO sits in on, the audits, the disaster-recovery program, the cost of running each tenant. That is this role.

You will report to the CTO and own SaaS operations end to end, from the Kubernetes clusters to the quarterly service review with a bank's executives. The mandate: make onboarding the fifth bank a checklist instead of a project.

What you'll own
  • Service management for every SaaS customer. SLA definition and reporting, incident communications and post-incident reviews, security questionnaires, audit cycles, and the day-to-day support model with each bank's service desk.
  • Release and deployment operations. The release calendar across the tenant estate, environment promotion and rollback discipline, production change management, and coordinating rollouts with each bank's change and freeze windows. Engineering builds the pipelines; you decide when and how production changes.
  • The reliability program. Tested disaster recovery and business continuity, backup verification, patching and vulnerability-management cadence, capacity planning, and a cost-per-tenant model the CFO can use.
  • The tenant onboarding runbook, so new banks go live on a repeatable path.
  • Our European launch. Two new Azure regions with data residency, DR, and support coverage in place, operated rather than merely deployed.
  • Compliance operations for outsourcing arrangements, working with bank risk teams under frameworks such as ISO 27001, SOC 2, and European and Gulf outsourcing regulation.
Your first six months
  • Four bank go-lives across two geographies, each with agreed SLAs, escalation paths, and incident procedures in place from day one.
  • A disaster-recovery exercise run and documented for at least one production environment.
  • An on-call rotation covering both regions without heroics.
  • A tenant cost model and a capacity plan for the whole estate.
What you bring
  • 12+ years in production operations, several of them leading the function for multi-tenant SaaS.
  • Ownership of customer-facing service management. You have led a severity-one bridge, presented the post-incident review to a customer's executives, and been through their audits.
  • You have built or scaled an SRE or platform-operations team and designed on-call across regions.
  • Enough Kubernetes and cloud depth to challenge your engineers on failure modes, DR design, and capacity claims.
  • Working knowledge of Azure itself. Regions and availability zones, networking and private connectivity, identity, and quota planning. You will negotiate these with Microsoft and with bank security teams.
  • Experience with certification and audit regimes: ISO 27001, SOC 2, or bank outsourcing regulation in the EU or the Gulf.
Nice to have
  • Wealth management, brokerage, or capital-markets domain exposure.
  • Experience operating streaming or workflow infrastructure (Temporal, Kafka, or similar).
  • Experience with on-premises software delivery to enterprise customers.
What they offer

Competitive compensation, a senior mandate reporting straight to the CTO, and a platform at the point where it goes from its first customers to a full production estate. The problems in this document are real and yours to fix.

By applying to this position, you are granting us permission to process your CV and keep your profile on file for consideration for this and future opportunities.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal SRE
Principal SRE

Alpheya • Abu Dhabi

On-site
AED 900,000 - 1,200,000
Lead SRE / Technology Operations (DevSecOps)
Lead SRE / Technology Operations (DevSecOps)

FinTop Consulting • Dubai

On-site
AED 450,000 - 750,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Epergne Solutions • Dubai

On-site
AED 200,000 - 300,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Epergne Solutions • Abu Dhabi

On-site
AED 180,000 - 250,000
Senior DevOps / Site Reliability Engineer (SRE)
Senior DevOps / Site Reliability Engineer (SRE)

Stellar Technologies • Abu Dhabi

On-site
AED 360,000 - 540,000
LEAD SOFTWARE ENGINEER - Azure DevOps
LEAD SOFTWARE ENGINEER - Azure DevOps

Happiest Minds Technologies • Abu Dhabi

On-site
AED 260,000 - 520,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AW Connect • Dubai

On-site
AED 420,000 - 660,000
Lead SRE / Technology Operations (DevSecOps)
Lead SRE / Technology Operations (DevSecOps)

Client of FinTop Consulting • Dubai

On-site
AED 420,000 - 640,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Epergne Solutions • Ras Al Khaimah

On-site
AED 469,000 - 603,000
Site Reliability Engineer - Digital Banking
Site Reliability Engineer - Digital Banking

DiceTek UAE • Al Ruways Industrial City

On-site