{bc}
indeed

Site Reliability Engineer (SRE)

DICETEK LLC
Abu Dhabi, UAE
Contract
Onsite
Discovered 2 days ago
Site Reliability EngineeringDevOpsSLIs, SLOs, and error budgetsObservabilityDynatracePrometheus
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

Site Reliability EngineeringDevOpsSLIs, SLOs, and error budgets
Smart Apply

Full Job Posting

What You Will Be Doing

  • Define and implement SLIs, SLOs, and error budgets for business-critical digital banking services.
  • Build observability across metrics, logs, traces, dashboards, and alerts using Dynatrace, Prometheus, Grafana, and ELK.
  • Use Dynatrace Davis AI or an equivalent AIOps platform for anomaly detection, predictive alerting, and proactive remediation.
  • Lead on-call triage, root-cause analysis, blameless postmortems, and actionable incident follow-ups.
  • Improve deployment safety with canary and blue-green deployments, rollout and rollback strategies, and readiness reviews.
  • Optimize microservices reliability, scalability, resilience, capacity, performance, and operational cost.
  • Automate runbooks, remediation scripts, proactive health checks, and self-healing workflows.
  • Embed reliability gates in CI/CD pipelines and maintain operational documentation, compliance, and security alignment.

Experience and Qualifications

  • 5+ years of SRE or DevOps experience with large-scale, high-availability systems.
  • Experience across banking, fintech, e-commerce, or other data-intensive digital ecosystems.
  • Bachelor’s degree in Computer Science or equivalent technical experience.
  • Strong Linux and performance troubleshooting experience.
  • Proven Terraform and Infrastructure as Code expertise.
  • Proficiency with Kubernetes and container orchestration in microservices environments.
  • AWS experience is preferred; Azure or GCP exposure is advantageous.
  • Experience with AIOps, anomaly detection, predictive alerting, and AI/ML-driven reliability automation.
  • Practical CI/CD experience with GitHub Actions, Jenkins, GitLab CI/CD, or Azure DevOps.
  • Experience with Kafka, RabbitMQ, Redis, Aurora, and RDS databases.
  • Strong scripting or programming skills in Python, Bash, or Go.

Ideal Candidate

  • Organized, structured, and meticulous in approach.
  • Experienced in cross-functional collaboration with distributed teams.
  • Analytical and skilled at troubleshooting complex production systems.
  • Calm and composed under pressure and capable of leading high-impact incidents.
  • Proactive in identifying issues and driving preventive improvements.
  • Passionate about AI-driven automation, observability, reliability engineering, cloud-native systems, and microservices.
  • Collaborative, adaptable, and comfortable working in a fast-paced, regulated environment.

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today

More from this employer

More jobs at DICETEK LLC