Principal SRE

Alpheya

United Arab Emirates

On-site

AED 600,000 - 1,000,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Alpheya, a wealth-management technology company in Abu Dhabi, seeks a senior SaaS operations leader to own service management end-to-end. You will govern SLAs, incident communications, disaster recovery, audits, and cost-per-tenant models across two geographies.

You will report to the CTO, work with bank customers, and drive a repeatable onboarding and release calendar for secure, compliant production environments.

Qualifications

  • 12+ years in production operations, multi-tenant SaaS experience.
  • Led severity-one bridges and post-incident reviews with executives.
  • Experience with ISO 27001, SOC 2, or EU/Gulf outsourcing regulations.
  • Deep Kubernetes and cloud expertise, DR design and capacity planning.
  • Familiar with streaming/workflow infrastructure such as Temporal or Kafka.

Responsibilities

  • Own SaaS operations end-to-end for all tenants and customers.
  • Define SLAs, report performance, and manage incident communications.
  • Lead audits, disaster recovery, backup verification, and cost per tenant.
  • Oversee release and deployment operations across environments.
  • Maintain an on-call rotation across regions and coordinate with teams.
  • Develop tenant onboarding runbooks for repeatable live deployments.

Skills

SRE leadership
Incident management
Kubernetes
Azure
On-call management
Cost modelling
DR planning
Audit readiness
Vendor coordination

Tools

Azure
Kubernetes
Temporal
Kafka

Job description

About Alpheya Alpheya is a wealth-management technology company headquartered in Abu Dhabi.

Banks are our customers.

We build each one a tailored investing experience, from mobile apps to advisor and back-office portals, on a single shared platform covering the full order-to-custody lifecycle.

The platform runs as SaaS on Microsoft Azure, with on-premises delivery for banks that require it.

The role Over the next six months we are taking multiple banks live as SaaS customers.

Our engineering teams build the platform and our SREs run the infrastructure.

What we don't yet have is a single leader accountable for the service our customers buy: the SLAs, the incident review a bank CIO sits in on, the audits, the disaster-recovery program, the cost of running each tenant.

That is this role.

You will report to the CTO and own SaaS operations end to end, from the Kubernetes clusters to the quarterly service review with a bank's executives.

The mandate: make onboarding the fifth bank a checklist instead of a project.

What you'll own Service management for every SaaS customer.

SLA definition and reporting, incident communications and post incident reviews, security questionnaires, audit cycles, and the day-to-day support model with each bank's service desk.

Release and deployment operations.

The release calendar across the tenant estate, environment promotion and rollback discipline, production change management, and coordinating rollouts with each bank's change and freeze windows.

Engineering builds the pipelines; you decide when and how production changes.

The reliability program.

Tested disaster recovery and business continuity, backup verification, patching and vulnerability-management cadence, capacity planning, and a cost-per-tenant model the CFO can use.

The tenant onboarding runbook, so new banks go live on a repeatable path.

Our European launch.

Two new Azure regions with data residency, DR, and support coverage in place, operated rather than merely deployed.

Compliance operations for outsourcing arrangements, working with bank risk teams under frameworks such as ISO 27001, SOC 2, and European and Gulf outsourcing regulation.

Your first six months Four bank go-lives across two geographies, each with agreed SLAs, escalation paths, and incident procedures in place from day one.

A disaster-recovery exercise run and documented for at least one production environment.

An on-call rotation covering both regions without heroics.

A tenant cost model and a capacity plan for the whole estate.

12+ years in production operations, several of them leading the function for multi-tenant SaaS.

Ownership of customer-facing service management.

You have led a severity-one bridge, presented the post-incident review to a customer's executives, and been through their audits.

You have built or scaled an SRE or platform-operations team and designed on-call across regions.

Enough Kubernetes and cloud depth to challenge your engineers on failure modes, DR design, and capacity claims.

Working knowledge of Azure itself.

Regions and availability zones, networking and private connectivity, identity, and quota planning.

You will negotiate these with Microsoft and with bank security teams.

Experience with certification and audit regimes: ISO 27001, SOC 2, or bank outsourcing regulation in the EU or the Gulf.

Nice to have Wealth management, brokerage, or capital-markets domain exposure.

Experience operating streaming or workflow infrastructure (Temporal, Kafka, or similar).

Experience with on-premises software delivery to enterprise customers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal SRE
Principal SRE

Alpheya • Abu Dhabi

On-site
AED 900,000 - 1,200,000
Principal SRE
Principal SRE

Professional.me • Abu Dhabi

On-site
AED 400,000 - 600,000
Senior SaaS Reliability & Service Operations Lead
Senior SaaS Reliability & Service Operations Lead

Alpheya • Abu Dhabi

On-site
AED 900,000 - 1,200,000
Senior SRE & SaaS Operations Leader
Senior SRE & SaaS Operations Leader

Professional.me • Abu Dhabi

On-site
AED 400,000 - 600,000
SaaS Reliability & Service Operations Lead
SaaS Reliability & Service Operations Lead

Alpheya • United Arab Emirates

On-site
AED 600,000 - 1,000,000
Site Reliability Engineer - Digital Banking
Site Reliability Engineer - Digital Banking

DiceTek UAE • Al Ruways Industrial City

On-site
Lead SRE / Technology Operations (DevSecOps)
Lead SRE / Technology Operations (DevSecOps)

Client of FinTop Consulting • Dubai

On-site
AED 420,000 - 640,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AW Connect • Dubai

On-site
AED 420,000 - 660,000
Lead SRE / Technology Operations (DevSecOps)
Lead SRE / Technology Operations (DevSecOps)

FinTop Consulting • Dubai

On-site
AED 450,000 - 750,000
Senior DevOps / Site Reliability Engineer (SRE)
Senior DevOps / Site Reliability Engineer (SRE)

Stellar Technologies • Abu Dhabi

On-site
AED 360,000 - 540,000