Base Career helps you apply smarter for this job.
Key skills for this role
Own disaster recovery architecture and execution across cloud environments.
Maintain DR solutions and backup strategy and run DR drills on a regular cadence; document gaps and drive remediation.
Design for resilience from the start and treat recoverability as a first-class requirement, not an afterthought.
Contribute to vulnerability management triage across infrastructure teams, threat detection, and infra security findings review.
Support PHI/PII classification scanning, penetration test coordination, and security exception approvals.
Maintain compliance controls; support HIPAA and SOC 2 audit readiness, access review and recertification, evidence collection, and change freeze coordination.
Maintain EKS container runtime security sensor coverage as part of ongoing platform hardening.
Design and manage roles and cross-account access controls following least-privilege principles across multi-account, multi-cloud environments.
Own cloud execution: load balancers, DNS (), VPC provisioning, and security group standards.
Manage compute resources at scale with an eye toward right-sizing and long-term maintainability.
Administer cloud storage and database services with attention to cost, performance, and resilience.
Own infrastructure health, cost, and performance monitoring using native and third-party tooling, building the observability that lets issues surface before they become incidents.
Administer and harden Linux and Windows Server environments across cloud and on-prem, including patching, performance tuning, troubleshooting, Active Directory integration, Group Policy, DNS, and certificate services.
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
, USA
Bengaluru, IND
, USA
, USA
Bengaluru, IND
Bengaluru, IND
, USA
, USA
, USA
Manage hybrid identity and authentication across on-prem and cloud workloads, and maintain OS-level security baselines and hardening standards across the estate.
Build and evolve CI/CD pipelines for secure, repeatable infrastructure deployments.
Write automation to reduce manual toil and enforce operational consistency across cloud and on-prem environments.
Take solutions from proof-of-concept to production with an eye toward long-term maintainability, not just getting it working once.
Monitor, scale, and maintain production infrastructure with availability, performance, and security as top priorities.
Participate in on-call rotation; serve as L2 escalation point for cross-team infrastructure support; lead root cause analysis and drive incident retrospectives to closure.
Author and maintain runbooks that hold up under pressure, not just at handoff.
Contribute to FinOps efforts: identify and remediate cost anomalies, own tagging remediation against enterprise tagging standards, and make pragmatic cost/performance/resilience tradeoffs.
Use AI coding assistants to accelerate IaC development, scripting, and troubleshooting.
Use AI tooling to draft first-pass runbooks, DR documentation, and incident retrospectives to validate and refine before publishing.
Document architecture, DR runbooks, and standard operating procedures others can actually follow under pressure.
Provide technical guidance to product teams on infrastructure resilience, migration sequencing, and recovery design.
Raises risk early rather than waiting for it to become an incident.
Comfortable with ambiguity in a large, multi-account, evolving cloud environment.
Strong sense of ownership; closes gaps rather than escalating and waiting.
Communicates technical tradeoffs clearly to both engineers and non-technical stakeholders.
Private healthcare technology company providing AI-powered software and services to health plans and health insurers.
Visit company websiteJobs and hiring trendsFull-time
Senior · 5+ years experience
Onsite
Apply faster on company sites with our extension.