Base Career helps you apply smarter for this job.
Key skills for this role
The Site Reliability Engineer (SRE) is responsible for the reliability, performance, and scalability of Precisely's infrastructure platforms across CEDAR (CCX) — an on-premises, private cloud managed services environment; RapidCX (RCX) — an AWS cloud environment for SaaS-delivered customer communications management; and Hosted Managed Services (HMS) — an AWS cloud environment supporting managed client deployments.
This role bridges software engineering and systems operations, building automation, observability tooling, and reliability standards to ensure platform availability and operational excellence. SREs are enabling partners: they set reliability standards, define what 'reliable' looks like for each service, validate production readiness, and coach engineering teams on operational best practices. Engineering teams own the reliability outcomes of the services they build; the SRE ensures they have the standards, tooling, and guidance to meet them.
As a Senior SRE, this role is regarded as a platform expert across CCX, RCX, and HMS, taking on the majority of complex reliability engineering work, mentoring less experienced engineers, and contributing to release triage and coordination efforts.
The Site Reliability Engineer (SRE) is responsible for the reliability, performance, and scalability of Precisely's infrastructure platforms across CEDAR (CCX) — an on-premises, private cloud managed services environment; RapidCX (RCX) — an AWS cloud environment for SaaS-delivered customer communications management; and Hosted Managed Services (HMS) — an AWS cloud environment supporting managed client deployments.
This role bridges software engineering and systems operations, building automation, observability tooling, and reliability standards to ensure platform availability and operational excellence. SREs are enabling partners: they set reliability standards, define what 'reliable' looks like for each service, validate production readiness, and coach engineering teams on operational best practices. Engineering teams own the reliability outcomes of the services they build; the SRE ensures they have the standards, tooling, and guidance to meet them.
As a Senior SRE, this role is regarded as a platform expert across CCX, RCX, and HMS, taking on the majority of complex reliability engineering work, mentoring less experienced engineers, and contributing to release triage and coordination efforts.
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
London, GBR
, AUS
, USA
, USA
, USA
, USA
London, GBR
, IND
, USA
Educational requirements (equivalent work experience will be accepted in place of the education requirement): Bachelor's degree in Computer Science, Information Systems, Engineering, or equivalent practical experience.
5+ years of systems or infrastructure engineering experience in an enterprise production environment.
High-level, developing subject-matter-expert competency in one or more complex infrastructure domains (e.g., cloud platform engineering, IaC automation, observability, or on-premises virtualization).
Advanced proficiency with Linux (RHEL/Oracle Linux) across multiple environments (on-premises and cloud, not just multi-site).
Proficient with Terraform for infrastructure-as-code; experienced with Ansible role development beyond basic hands-on use.
Experience deploying and managing workloads in AWS at an intermediate level: EC2, ECS, S3, VPC, IAM, CloudWatch, Auto Scaling.
Proficiency with at least one scripting language (Python, Bash) for automation development.
Demonstrated experience designing monitoring system architecture and alerting strategy — not just operating existing dashboards (Datadog preferred).
Solid understanding of TCP/IP networking, DNS, load balancing, and distributed systems.
Experience designing and managing CI/CD pipelines and deployment automation standards.
Strong analytical skills; demonstrated experience authoring root cause analyses for complex incidents and identifying systemic, preventive fixes from recurring incident patterns.
Ability to define SLOs and lead Operational Readiness Reviews (ORRs); comfortable partnering with engineering teams on production readiness.
Demonstrated ability to work cross-functionally with engineering teams on reliability standards and observability requirements.
Experience participating in change advisory processes and contributing to capacity and reliability planning.
Travel is required: No — approximately 0%.
Precisely is a data integrity software company serving enterprises with software, data, and data strategy consulting services.
Visit company websiteJobs and hiring trendsFull-time
Senior · 5+ years experience
Remote
Apply faster on company sites with our extension.