{bc}
adp

Cloud Platforms Engineer

ECS FEDERAL LLC
Fairfax, USA
Full-time
Senior · 5+ years experience
Hybrid
USD 180000-210000 yearly / year
Discovered 1 weeks ago
TerraformInfrastructure as Code (IaC)Site Reliability EngineeringDevOps
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

TerraformInfrastructure as Code (IaC)Site Reliability Engineering
Smart Apply

Full Job Posting

Required Skills

  • U. S. Citizen.
  • Active DoD Secret security clearance.
  • Bachelor’s degree with 5+ years of related work experience.
  • Ability to obtain a DoD 8140 IAT Level II Security+ (or higher) within 60 days of hire.
  • Ability to work in a hybrid capacity in Fairfax, VA (up to 3 days in office).
  • Ability to travel <20% throughout the lifespan of the Program to CONUS / OCONUS customer sites and government installations.
  • Strong experience designing, deploying, and supporting highly available, fault-tolerant cloud infrastructure.
  • Hands-on experience with Terraform and Infrastructure as Code (IaC) principles for provisioning and managing cloud resources.
  • Knowledge of High Availability (HA) and Disaster Recovery (DR) architecture, including redundancy, automated failover, backup and restoration, geographic resiliency, and recovery planning.
  • Experience managing the full cloud infrastructure lifecycle, including provisioning, configuration, patching, upgrades, vulnerability remediation, maintenance, and decommissioning.
  • Strong infrastructure automation skills with an emphasis on reducing manual administration, configuration drift, and operational error.
  • Experience implementing and operating monitoring, logging, alerting, and observability solutions for production infrastructure.
  • Strong troubleshooting and diagnostic skills, including incident response, root-cause analysis, and corrective action development.
  • Understanding of Site Reliability Engineering concepts, including SLIs, SLOs, availability targets, RTOs, and RPOs.
  • Experience developing and validating disaster recovery procedures, including failover exercises, recovery testing, and infrastructure restoration.
  • Ability to evaluate infrastructure capacity, performance, scalability, availability, and resiliency and recommend architectural or operational improvements.
  • Experience working collaboratively with application development, cybersecurity, DevOps, and platform engineering teams.
  • Ability to develop and maintain technical documentation, including architecture diagrams, operational procedures, runbooks, and recovery documentation.
  • Strong understanding of cloud infrastructure security, vulnerability management, and operational best practices.
  • Demonstrated ability to provide technical leadership and guidance in infrastructure reliability, resiliency, automation, and cloud operations.
  • Experience supporting production, enterprise, regulated, or mission-critical environments is highly desirable.
  • Strong problem-solving and decision-making capabilities, with a proven ability to weigh the relative costs and benefits of potential actions and identify the most appropriate solution.
  • Highly developed interpersonal and oral/written communication skills, with the ability to effectively and professionally interact with a diverse set of stakeholders (from peers to end-users to executive management).

Desired Skills

  • Bachelor’s degree in Computer Science, Information Technology, Computer Networking, Computer Security, or other Science, Technology, Engineering and Mathematics (STEM) discipline.
  • Demonstrated experience administering and engineering cloud infrastructure in production environments.
  • Strong hands-on experience with Terraform and Infrastructure as Code (IaC) practices.
  • Experience developing and maintaining infrastructure automation for provisioning and operational tasks.
  • Knowledge of High Availability (HA) and Disaster Recovery (DR) architecture and implementation.
  • Experience with monitoring, logging, alerting, and observability solutions.
  • Strong operational troubleshooting, incident response, and root-cause analysis skills.
  • Experience supporting enterprise-scale, regulated, or government environments.
  • Familiarity with security, compliance, vulnerability management, and governance requirements in cloud environments.
  • Ability to work effectively across infrastructure, cybersecurity, DevOps, platform engineering, and application teams.
  • Strong understanding of cloud reliability, resiliency, scalability, and operational best practices.

#EverforthECS1

ECS Federal LLC is an equal opportunity employer and does not discriminate or allow discrimination on the basis any characteristic protected by law. All qualified applicants will receive consideration for employment without regard to disability, status as a protected veteran or any other status protected by applicable federal, state, or local jurisdiction law.

Everforth ECS is the federal segment of Everforth , a $4B global organization with over 10,000 employees. Our nearly 3,500 professionals deliver advanced technology solutions in data and AI, cybersecurity, and enterprise transformation, serving defense, intelligence, and federal civilian agencies. Our work powers mission-critical outcomes, strengthens technology partnerships, and creates meaningful opportunities for our people. We are defined by a commitment to excellence in delivery, a culture of innovation, and an environment where talent can thrive and grow. We value:

Attracting and developing top talent and high-performing teams

Fostering a culture that is engaging, accountable, and mission-driven

Meet the challenge. Make a difference with Everforth ECS!

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today

More from this employer

More jobs at ECS FEDERAL LLC