Site Reliability Engineer (SRE)
Job Fit Check
Base Career helps you apply smarter for this job.
Key skills for this role
Role Overview
The Site Reliability Engineer will ensure the reliability, availability, performance, and scalability of critical applications and infrastructure.
The role requires experience in monitoring, automation, cloud technologies, incident management, and DevOps practices.
Key Skills for This Role
Full Job Posting
Job purpose
The Site Reliability Engineer will ensure the reliability, availability, performance, and scalability of critical applications and infrastructure.
The role requires experience in monitoring, automation, cloud technologies, incident management, and DevOps practices.
Key responsibilities
- Monitor and maintain application and infrastructure availability, performance, and reliability.
- Design and implement monitoring, logging, and alerting solutions.
- Manage production incidents and perform root cause analysis.
- Automate operational and deployment processes.
- Collaborate with development, infrastructure, and DevOps teams to improve performance and resilience.
- Implement and maintain CI/CD pipelines and Infrastructure as Code.
- Support containerized environments and cloud infrastructure.
- Develop scripts and automation tools to reduce manual operational work.
- Implement observability solutions including monitoring, logging, and distributed tracing.
- Document operational procedures, incidents, and system configurations.
Technical skills
- Monitoring tools include Prometheus, Grafana, and Zabbix.
- Logging tools include ELK or Elastic Stack and Splunk.
- APM tools include Dynatrace, AppDynamics, and New Relic.
- Cloud platforms include AWS, Microsoft Azure, and GCP.
- Container technologies include Docker, Kubernetes, and OpenShift.
- CI/CD tools include Jenkins, GitLab CI/CD, and Azure DevOps.
- Infrastructure as Code tools include Terraform and Ansible.
- Version control tools include Git, GitHub, and GitLab.
- Incident management tools include ServiceNow and PagerDuty.
- Distributed tracing tools include OpenTelemetry and Jaeger.
- Scripting languages include Bash, Python, and PowerShell.
- Relevant database and web technologies include PostgreSQL, Oracle, SQL Server, IIS, Nginx, Apache, and REST APIs.
Qualifications and experience
- A bachelor's degree in Computer Science, Information Technology, or a related field.
- Proven experience as a Site Reliability Engineer, DevOps Engineer, or Production Support Engineer.
- Strong experience in cloud infrastructure, automation, monitoring, and incident management.
- Hands-on experience with Kubernetes and containerized environments.
- Strong troubleshooting and root cause analysis skills.
- Experience in highly available, large-scale production environments.
- Excellent communication and collaboration skills.
Apply for this job in 1 click
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
More jobs at DICETEK LLC
Tele caller
, UAE
Job Purpose: We are looking for energetic and customer-focused Telecallers to handle outbound and inbound calls for banking-related products and services. The role involves engaging with customers, understanding their re
Project Manager-Finance and Risk
, UAE
Experience /Skills Required Experience in managing Banking projects in general. Specific experience in managing Finance / Risk projects is a plus Exposure to banking or financial regulatory environments. JIRA Expertis
SRE (Site Reliability Engineer)
, UAE
SRE (Site Reliability Engineer): Responsible for ensuring application and infrastructure reliability, availability, performance, and operational efficiency through monitoring, automation, and incident management. Require
Head Of Artificial Intelligence and Advanced Analytics
, UAE
Minimum Qualifications and Experience: PhD in Computer Engineering - 0 PhD in Information Technology - 0 MSc in Computer Engineering - 3 MSc in Information Technology - 3 BSc in Information Technology - 6 BSc in Computer
Director – Policies and Studies Department (UAE National only)
, UAE
Director – Policies and Studies Department Nationality: UAE National (Male or Female) Work Location: Dubai Experience: Minimum 10 years of experience, including 3–5 years in a managerial role or as a Head of Section i
System Lead – Enterprise Fraud Risk Management Systems
, UAE
Job Summary We are seeking an experienced Systems Manager to lead the management, enhancement, and ongoing support of the Bank's enterprise fraud risk management applications Clari-5, ensuring they effectively support th
Senior Inspector – Public Health (UAE Nationals)
, UAE
Qualification required: Bachelor’s degree in public health or any related field Experience required : 3-7 years of experience Nationality: Emirati with Family Book ( ONLY) Preference UAE Nationals from Dubai (Hatta), S
Lead, Fraud Analytics and MI
, UAE
Role Purpose To lead the bank's Fraud Analytics, Rules Management and Reporting function. The role owns end-to-end fraud rules lifecycle on Clari5 EFM and HPS Power CARD, drives rule performance optimization, gap analysi
Tele caller
, UAE
Project Manager-Finance and Risk
, UAE
SRE (Site Reliability Engineer)
, UAE
Head Of Artificial Intelligence and Advanced Analytics
, UAE
Director – Policies and Studies Department (UAE National only)
, UAE
System Lead – Enterprise Fraud Risk Management Systems
, UAE
Senior Inspector – Public Health (UAE Nationals)
, UAE
Lead, Fraud Analytics and MI
, UAE