{bc}
linkedin

Site Reliability Engineer (SRE)

Dicetek LLC
Dubai, UAE
Contract
Mid-Senior
Onsite
Discovered 1 weeks ago
Site reliability engineeringMonitoring and alertingPrometheusGrafanaElastic StackSplunk
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

Site reliability engineeringMonitoring and alertingPrometheus
Smart Apply

Full Job Posting

Job Purpose

The Site Reliability Engineer ensures the reliability, availability, performance, and scalability of critical applications and infrastructure.

The role requires experience with monitoring, automation, cloud technologies, incident management, and DevOps practices.

Key Responsibilities

  • Monitor and maintain application and infrastructure availability, performance, and reliability.
  • Design monitoring, logging, and alerting solutions and manage production incidents through root cause analysis.
  • Automate operational and deployment processes and maintain CI/CD pipelines and infrastructure as code.
  • Work with development, infrastructure, and DevOps teams to improve application performance and resilience.
  • Support containers and cloud infrastructure and develop scripts to reduce manual operations.
  • Implement observability with monitoring, logging, and distributed tracing.
  • Document operational procedures, incidents, and system configurations.

Required Technical Skills

  • Monitoring tools include Prometheus, Grafana, and Zabbix; logging tools include ELK or Elastic Stack and Splunk.
  • APM tools include Dynatrace, AppDynamics, and New Relic.
  • Cloud experience may include AWS, Microsoft Azure, or GCP.
  • Container experience includes Docker, Kubernetes, and OpenShift.
  • CI/CD tools include Jenkins, GitLab CI/CD, and Azure DevOps.
  • Infrastructure as code tools include Terraform and Ansible; version control includes Git, GitHub, and GitLab.
  • Incident management tools include ServiceNow and PagerDuty; distributed tracing includes OpenTelemetry and Jaeger.
  • Scripting may use Bash, Python, or PowerShell, with database experience in PostgreSQL, Oracle, or SQL Server.
  • Web and API technologies include IIS, Nginx, Apache, and REST APIs.

Qualifications and Experience

  • A bachelor's degree in Computer Science, Information Technology, or a related field is required.
  • Proven experience as a Site Reliability Engineer, DevOps Engineer, or Production Support Engineer is required.
  • Strong experience in cloud infrastructure, automation, monitoring, and incident management is required.
  • Hands-on experience with Kubernetes and containerized environments is required.
  • Strong troubleshooting, root cause analysis, communication, and collaboration skills are required.
  • Experience in highly available, large-scale production environments is required.

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today

More from this employer

More jobs at Dicetek LLC