{bc}
linkedin

Site Reliability Engineer - full stack and Grafana

Wareef
Riyadh, KSA
Full-time
Entry
Hybrid
Discovered 4 weeks ago
Site Reliability EngineeringFull stack technologiesGrafanaPrometheusLinuxNetworking fundamentals
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

Site Reliability EngineeringFull stack technologiesGrafana
Smart Apply

Full Job Posting

About the Role

This is a full-time, on-site role based in Riyadh for a Site Reliability Engineer with full stack and Grafana expertise.

The role involves designing, implementing, and maintaining reliable, scalable, and secure systems across application and infrastructure layers.

Daily responsibilities include monitoring system performance, building and maintaining dashboards in Grafana, proactively identifying issues, and troubleshooting incidents to reduce downtime and improve service quality.

The engineer will collaborate closely with software development, operations, and infrastructure teams to automate workflows, optimize deployments, and enhance observability and alerting.

The role also includes participating in on-call rotations, contributing to incident response and post-incident reviews, and continuously refining reliability best practices and standards.

Responsibilities

  • Design, implement, and maintain reliable, scalable, and secure systems across application and infrastructure layers.
  • Monitor system performance and build and maintain dashboards in Grafana.
  • Proactively identify issues and troubleshoot incidents to reduce downtime and improve service quality.
  • Collaborate closely with software development, operations, and infrastructure teams to automate workflows, optimize deployments, and enhance observability and alerting.
  • Participate in on-call rotations, contribute to incident response and post-incident reviews, and continuously refine reliability best practices and standards.

Qualifications

  • Site Reliability Engineering skills, including designing for reliability, scalability, observability, and incident response.
  • Strong troubleshooting and problem-solving skills for application, infrastructure, and performance issues.
  • Software development experience with full stack technologies (e.g., backend services, APIs, frontend frameworks) and scripting for automation.
  • System administration capabilities, including managing Linux-based environments, networking fundamentals, and security best practices.
  • Infrastructure skills in deploying, operating, and optimizing cloud or hybrid environments, CI/CD pipelines, and containerization/orchestration tools.
  • Hands-on experience with monitoring and visualization tools such as Grafana, Prometheus, or similar observability platforms.
  • Bachelor’s degree in Computer Science, Engineering, Information Technology, or a related field, or equivalent practical experience.
  • Ability to work collaboratively in cross-functional teams, communicate clearly, and document processes and runbooks effectively.
  • Experience with automation frameworks, Infrastructure as Code (e.g., Terraform, Ansible), and version control systems is beneficial.
  • Previous experience in large-scale or mission-critical environments and familiarity with ITIL or similar service management frameworks is a plus.

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today