Site Reliability Engineer III
Job Fit Check
Base Career helps you apply smarter for this job.
Key skills for this role
Role Overview
Veeam is hiring a Site Reliability Engineer III to serve as a hands-on technical leader within its SRE team.
The role guides engineers, influences product development, and ensures systems are reliable, scalable, and observable.
The position drives strategic initiatives, mentoring, and architectural best practices across the platform.
Key Skills for This Role
Full Job Posting
Role overview
Veeam is hiring a Site Reliability Engineer III to serve as a hands-on technical leader within its SRE team.
The role guides engineers, influences product development, and ensures systems are reliable, scalable, and observable.
The position drives strategic initiatives, mentoring, and architectural best practices across the platform.
What you'll do
- Design and evolve highly available, fault-tolerant, and scalable infrastructure across public clouds, initially Azure.
- Establish SLIs, SLOs, and error budgets for reliability objectives.
- Lead incident response, analysis, blameless postmortems, and knowledge-sharing sessions.
- Drive comprehensive telemetry, logs, metrics, and tracing practices.
- Develop automation and self-healing tools to reduce operational toil.
- Participate in on-call rotations and lead operational excellence.
- Contribute to IaC, CI/CD, deployment automation, and configuration management.
- Integrate monitoring and chaos engineering tools to test reliability under load and failure.
- Implement testing, canary deployments, and release validation pipelines.
- Mentor engineers and promote DevOps and SRE practices across global teams.
What you'll bring
- 5+ years of hands-on Software Engineering experience, including at least 2 years in SRE, Platform Engineering, or a similar field.
- Deep experience building systems on public cloud providers, with Azure preferred.
- Strong programming skills in JavaScript, Node.js, TypeScript, Go, Java, C#, or a similar language.
- Experience delivering monitoring, alerting, and observability tooling such as Prometheus, Grafana, or OpenTelemetry.
- Experience with Terraform or Pulumi and container orchestration such as Kubernetes.
- Understanding of distributed systems, cloud networking, and cloud-native system design.
- Excellent communication and collaboration skills across geographies and disciplines.
Bonus skills
- Experience with large-scale B2B SaaS platforms.
- Background in chaos engineering, resilience testing, performance testing, load testing, or incident learning programs.
- Familiarity with compliance frameworks such as ISO, SOC 2, GDPR, FEDRAMP, or CMMC.
Benefits and development
- The posting lists paid vacation, global wellbeing days, volunteer hours, medical coverage, insurance, wellbeing support, counselling, transportation benefits, and learning opportunities.
Location eligibility
- Veeam states that applicants permanently located outside India may be declined.
Apply for this job in 1 click
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
More jobs at Veeam Software
Talent Growth & Program Manager (REMOTE WEST COAST)
, USA
Revenue Intelligence Operations Manager, Marketing
, USA
Staff Software Development Engineer in Test
San Jose, USA
Country Manager - Turkey
New Providence, USA
Senior Engineer - Hyperscale Analytics
San Jose, USA
Account Executive, Commercial Accounts
Mexico, USA
Senior Engineer - Access Entitlements
Columbia, CAN
Distribution Sales Account Manager
Brazil, USA