Senior Site Reliability Engineer, CloudOps
Job Fit Check
Base Career helps you apply smarter for this job.
Role Overview
We are seeking a Senior Site Reliability Engineer, CloudOps to support, scale, and optimize a multi-account AWS environment hosting healthcare-oriented applications and analytics platforms.
In this role, you will bridge infrastructure engineering, operational reliability, production support, and cloud modernization initiatives across complex microservices architectures.
The ideal candidate brings strong AWS expertise, solid Linux administration skills, and a proven track record of managing production systems in HIPAA/HiTrust regulated environments.
You will participate in incident response, on-call rotations, and continuous deployment workflows while helping drive our transition toward containerized and Kubernetes-based platforms.
This collaborative position is built for an analytical engineer who excels at resolving production incidents, partnering with developers, and continuously elevating operational excellence.
Full Job Posting
Position Summary
We are seeking a Senior Site Reliability Engineer, CloudOps to support, scale, and optimize a multi-account AWS environment hosting healthcare-oriented applications and analytics platforms.
In this role, you will bridge infrastructure engineering, operational reliability, production support, and cloud modernization initiatives across complex microservices architectures.
The ideal candidate brings strong AWS expertise, solid Linux administration skills, and a proven track record of managing production systems in HIPAA/HiTrust regulated environments.
You will participate in incident response, on-call rotations, and continuous deployment workflows while helping drive our transition toward containerized and Kubernetes-based platforms.
This collaborative position is built for an analytical engineer who excels at resolving production incidents, partnering with developers, and continuously elevating operational excellence.
Essential Duties & Responsibilities
- Manage, maintain, and troubleshoot a multi-account AWS Organization environment (35+ accounts) and core services, including EC2, ECS/Fargate, Lambda, S3, CloudFront, API Gateway, and Aurora/RDS databases.
- Support production deployments, CI/CD pipelines (Jenkins, AWS CodePipeline), and infrastructure automation using Python, Bash, and AWS CloudFormation.
- Monitor system health and performance using Datadog, CloudWatch, and Zabbix; investigate alerts, execute root-cause analysis, and refine monitoring coverage to reduce operational noise.
- Participate in a shared on-call rotation, managing incident response and performing failover/recovery validation for production applications and data stores.
- Maintain HIPAA/HiTrust compliance and security posture by managing tools like Prisma/Cortex Cloud, Security Hub, and GuardDuty, while enforcing proper IAM policies and network segmentation.
- Support Java (Spring Boot) and Python applications running in containers, assisting developers during investigations and preparing for future Kubernetes (EKS) modernization initiatives.
Knowledge & Skills
- Deep hands-on expertise with AWS core services (networking, compute, serverless, and database technologies) and CloudFormation IaC automation.
- Strong Linux administration skills (primarily Ubuntu) along with proficiency in Python and Bash scripting for operational automation.
- Experience with containerization technologies (Docker, ECS/Fargate) and familiarity with modern Kubernetes ecosystems (EKS, Helm, ArgoCD).
- Solid understanding of observability tools (Datadog, CloudWatch, Zabbix) and CI/CD pipelines (Jenkins, CodePipeline, Git workflows).
- Knowledge of cloud security best practices, access management (IAM), and compliance frameworks within regulated sectors (HIPAA/HiTrust).
- Proven diagnostic, incident-management, and analytical troubleshooting skills for complex microservices architectures.
Minimum Qualifications, Education & Experience
- Must be at least 18 years of age.
- High School Diploma required.
- Bachelor’s degree from an accredited college or university is required.
- 7+ years of hands-on experience in AWS Cloud Engineering, DevOps, Site Reliability Engineering (SRE), or Infrastructure Engineering.
- Practical background supporting production workloads in Linux/AWS environments, reading application logs, and making minor code fixes.
- Direct experience participating in on-call rotations and incident response protocols.
- Prior experience in the healthcare industry maintaining HIPAA/HiTrust-compliant infrastructure.
Work Environment
- This is largely a sedentary role.
- This job operates in a professional office environment and routinely uses standard office equipment.
- Typically requires travel less than 5% of the time
About ICU Med Careers
Develops and sells infusion and critical-care medical products.
Visit company websiteApply for this job in 1 click
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
More jobs at ICU Med Careers
Technician, Spare Parts - 2nd Shift
Southington, USA
Position Summary The primary function of the spare parts clerk is to procure, record, store repair parts and supplies. Work is performed under the general direction of the Spare Parts Administrator. Essential Duties & Re
Senior QA Inspector
Southington, USA
Position Summary The Sr. QA Inspector performs testing and inspection of sub-assemblies, final product, commodities, GMP, BOP, safety compliance audits. The Sr. QA Inspector enforces GMP compliance as outlined in interna
Inspector II
Southington, USA
Position Summary The Inspector II will perform moderately complex, repetitive tasks associated with quality inspections. The incumbent will ensure company products are in compliance with internal and external specificati
Engineer III, Molding
Salt Lake City, USA
Position Summary The Engineer III, Molding is responsible for supporting injection molding improvements through adherence to established product design, scientific injection molding process, tooling, resin, and equipment
Mold Technician II - 2nd Shift
Southington, USA
Position Summary The Mold Technician performs routine maintenance on plastic injection molds both in the press and on the bench under the direct supervision of a Lead Technician or Senior Mold Technician. This includes d
Device Sales Specialist - Southeast
, USA
Position Summary The Device Sales Specialist is responsible for building and maintaining relationships with key decision makers that lead to future business opportunities. This position increases profitability and expand
Analyst, Transportation - International
, USA
Position Summary The Transportation Analyst is responsible for planning and monitoring shipments within ICU’s distribution network, ensuring a high level of customer service with a focus on managing costs. This individua
Lead, Warehouse
Southington, USA
Position Summary The Warehouse Lead will be responsible for overseeing the daily warehouse responsibilities of all areas assigned. The Warehouse Lead will direct employees in the following tasks: picking and staging cust
Technician, Spare Parts - 2nd Shift
Southington, USA
Senior QA Inspector
Southington, USA
Inspector II
Southington, USA
Engineer III, Molding
Salt Lake City, USA
Mold Technician II - 2nd Shift
Southington, USA
Device Sales Specialist - Southeast
, USA
Analyst, Transportation - International
, USA
Lead, Warehouse
Southington, USA