{bc}
linkedin

Systems Engineer - Infrastructure & Disaster Recovery (National Talent)

Dubai Department of Economy and Tourism
Dubai, UAE
Full-time
Mid-Senior
Onsite
Discovered 1 weeks ago
Windows ServerEnterprise LinuxActive DirectoryGroup PolicyDNS and DHCPVirtualization platforms and clusters
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

Windows ServerEnterprise LinuxActive Directory
Smart Apply

Full Job Posting

Job Objective

The Systems Engineer, Infrastructure & Disaster Recovery maintains, optimizes, and recovers server, virtualization, container, backup, and selected cloud infrastructure.

The role keeps infrastructure services available, secure, supportable, and recoverable through administration, monitoring, automation, backup management, and disaster-recovery practices.

The role resolves complex incidents and works with Network, Database, Applications, DevOps, and Cybersecurity teams.

Infrastructure Platform Administration

  • Administer supported Windows Server and enterprise Linux environments.
  • Provision, configure, patch, upgrade, and decommission physical and virtual servers.
  • Administer Active Directory, Group Policy, DNS, DHCP, file services, and assigned infrastructure services.
  • Maintain security hardening and supported-version standards.
  • Monitor availability, performance, capacity, and resource utilization and resolve complex incidents.
  • Support server migration, consolidation, lifecycle renewal, and technology refresh activities.

Virtualization and Containers

  • Administer approved virtualization platforms and clusters and provision, resize, migrate, and decommission virtual machines.
  • Maintain high availability, resource balancing, live migration, and host, cluster, and virtual-machine monitoring.
  • Administer Kubernetes cluster and node lifecycle, access controls, certificates, platform health, high availability, backups, and restoration.
  • Troubleshoot control-plane, node, scheduling, capacity, and cluster-level incidents.
  • Coordinate application deployments and workload requirements with DevOps and Application teams.

Backup and Data Protection

  • Administer enterprise backup and recovery platforms for servers, virtual machines, file systems, infrastructure services, and assigned components.
  • Maintain backup schedules, retention requirements, storage policies, replication, and recovery copies.
  • Monitor backup, replication, and recovery activities and resolve failed or incomplete jobs.
  • Conduct restoration tests and maintain secure, immutable, or isolated recovery copies.
  • Identify coverage, capacity, and recoverability gaps and maintain restoration procedures and reports.

Disaster Recovery

  • Maintain systems-domain disaster-recovery procedures, recovery sequences, and technical runbooks.
  • Align recovery arrangements with approved recovery priorities, Recovery Time Objectives, and Recovery Point Objectives.
  • Maintain dependencies and recovery procedures for servers, virtual machines, infrastructure services, Kubernetes, and selected cloud workloads.
  • Execute restoration, failover, failback, and disaster-recovery exercises and validate post-recovery health.
  • Record recovery results and gaps, recommend improvements, and maintain patched, operational recovery environments.
  • Reflect production changes in recovery configurations, capacity requirements, and runbooks.

Cloud and Automation

  • Administer selected cloud virtual machines, storage, and infrastructure-recovery services.
  • Conduct approved cloud recovery tests, failovers, reprotection, and failback activities.
  • Monitor cloud infrastructure availability, performance, and resource utilization.
  • Automate provisioning, configuration, patching, monitoring, and recovery activities.
  • Maintain scripts, templates, and configuration files in approved version-control repositories.
  • Apply testing, peer review, security validation, and change control to automated infrastructure changes.
  • Identify configuration drift and repetitive administrative activities suitable for automation.
  • The role does not own enterprise cloud architecture, application deployment pipelines, or application-release automation.

Requirements

  • Ability to administer Windows Server and enterprise Linux environments.
  • Ability to manage virtualization platforms, clusters, virtual machines, and high-availability capabilities.
  • Ability to administer Kubernetes infrastructure, including lifecycle management, access controls, certificates, and cluster restoration.
  • Ability to manage enterprise backup, restoration testing, retention, recovery copies, and recoverability gaps.
  • Ability to maintain disaster-recovery procedures, runbooks, recovery sequences, and recovery environments.
  • Ability to automate infrastructure provisioning, configuration, patching, monitoring, and recovery using scripting and infrastructure-as-code tools.
  • Ability to apply testing, peer review, security validation, and change control to automated infrastructure changes.
  • Ability to coordinate infrastructure work with Network, Database, Applications, DevOps, and Cybersecurity teams.

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today

More from this employer

More jobs at Dubai Department of Economy and Tourism