{bc}
linkedin

Principal SRE

Alpheya
Abu Dhabi, UAE
Full-time
Director
Onsite
Discovered 1 weeks ago
SaaS operationsService management and SLA managementIncident management and post-incident reviewsRelease and deployment operationsKubernetesMicrosoft Azure
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

SaaS operationsService management and SLA managementIncident management and post-incident reviews
Smart Apply

Full Job Posting

About Alpheya

Alpheya is a wealth-management technology company headquartered in Abu Dhabi.

The company provides banks with tailored investing experiences across mobile, advisor, and back-office portals.

Its shared platform covers the order-to-custody lifecycle and runs as SaaS on Microsoft Azure, with on-premises delivery where required.

The Role

The Principal SRE will own the customer service delivered by the SaaS platform, including SLAs, incident reviews, audits, disaster recovery, and tenant operating costs.

The role reports to the CTO and owns SaaS operations end to end, from Kubernetes clusters to quarterly service reviews with bank executives.

The mandate is to make onboarding additional banks a repeatable operational checklist rather than a project.

Core Ownership Areas

  • Own SLA definition and reporting, incident communications, post-incident reviews, security questionnaires, audit cycles, and customer service-desk support models.
  • Own release calendars, environment promotion, rollback discipline, production change management, and coordination with bank change and freeze windows.
  • Lead disaster recovery, business continuity, backup verification, patching, vulnerability management, capacity planning, and tenant cost modeling.
  • Create a repeatable tenant onboarding runbook for new bank go-lives.
  • Operate two new Azure regions with data residency, disaster recovery, and support coverage.
  • Manage outsourcing compliance operations with bank risk teams under ISO 27001, SOC 2, and European or Gulf outsourcing regulation.

First Six Months

  • Deliver four bank go-lives across two geographies with agreed SLAs, escalation paths, and incident procedures.
  • Run and document a disaster-recovery exercise for at least one production environment.
  • Establish an on-call rotation covering both regions without relying on exceptional individual effort.
  • Create a tenant cost model and capacity plan for the full estate.

Required Experience

  • At least 12 years in production operations, including several years leading multi-tenant SaaS operations.
  • Experience owning customer-facing service management, leading severity-one bridges, presenting post-incident reviews to customer executives, and completing customer audits.
  • Experience building or scaling an SRE or platform-operations team and designing regional on-call coverage.
  • Kubernetes and cloud expertise sufficient to evaluate failure modes, disaster recovery design, and capacity claims.
  • Working knowledge of Azure regions, availability zones, networking, private connectivity, identity, and quota planning.
  • Experience with ISO 27001, SOC 2, or bank outsourcing regulation in the European Union or Gulf.

Preferred Experience

  • Wealth management, brokerage, or capital-markets domain exposure is preferred.
  • Experience operating streaming or workflow infrastructure such as Temporal or Kafka is preferred.
  • Experience delivering on-premises software to enterprise customers is preferred.

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today

More from this employer

More jobs at Alpheya