{bc}
indeed

Senior Site Reliability Engineer

Okta
Karnataka, IND
Hybrid
Discovered 6 days ago
Cloud infrastructureAWS or GCPKubernetesTerraformHelmGolang or Python
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

Cloud infrastructureAWS or GCPKubernetes
Smart Apply

Full Job Posting

Engineering Opportunity

Okta is hiring an experienced Senior Site Reliability Engineer for its Emerging Products Group.

The team builds reliable, scalable, and secure cloud services using automation, platform engineering, observability, and operational excellence.

The role partners with software engineers, architects, and product teams to design, build, and operate cloud services.

Reliability and Operations

  • Design, build, and operate large-scale cloud infrastructure and production services.
  • Participate in an on-call rotation supporting highly available customer-facing systems.
  • Lead incident response and post-incident reviews focused on systemic improvements.
  • Define and improve SLIs, SLOs, and error budgets.
  • Improve service availability, scalability, performance, and resilience.
  • Improve observability through metrics, logging, tracing, dashboards, and alerting.

Engineering and Automation

  • Develop software, automation, and infrastructure using Go, Python, Terraform, and related technologies.
  • Eliminate operational toil through automation, tooling, and platform engineering.
  • Improve deployment safety and workflows through CI/CD and GitOps practices.
  • Modernize existing workloads and align them with evolving platform capabilities.
  • Build self-service platforms, operational guardrails, and automation for developer velocity, reliability, and security.

Technical Excellence

  • Strong experience operating large-scale production services in AWS and/or GCP.
  • Deep production expertise with Kubernetes, including networking, storage, scheduling, scaling, and workload lifecycle troubleshooting.
  • Extensive experience with Infrastructure as Code technologies such as Terraform and Helm.
  • Strong software engineering skills in Golang and/or Python.
  • Experience building automation and internal engineering platforms.
  • Experience operating distributed data platforms such as PostgreSQL, Redis, OpenSearch, MySQL, Cassandra, or similar technologies.
  • Strong understanding of cloud networking, including DNS, load balancing, ingress, TLS, service networking, and traffic management.
  • Experience with observability platforms, monitoring strategies, and production telemetry.
  • Experience with or strong interest in AI-assisted engineering and operational automation.

Operational Excellence

  • Strong expertise operating customer-facing production systems.
  • Experience leading incident response and driving operational improvements.
  • Deep understanding of SLIs, SLOs, error budgets, and capacity planning.
  • Strong understanding of CI/CD pipelines, deployment strategies, and automation-first operations.
  • Ability to balance reliability, scalability, security, and engineering velocity.

Workplace

  • The role is associated with a hybrid work arrangement.
  • Okta provides an immersive, in-person onboarding experience.

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today

More from this employer

More jobs at Okta