{bc}
indeed

Senior Site Reliability Engineer

Okta
Karnataka, IND
Onsite
Discovered 6 days ago
Site Reliability EngineeringCloud infrastructureAWSGoogle Cloud PlatformKubernetesTerraform
Free

Job Fit Check

Base Career helps you apply smarter for this job.

?%
Ready to Scan

Key skills for this role

Site Reliability EngineeringCloud infrastructureAWS
Smart Apply

Full Job Posting

About Okta

Okta provides Workforce and Customer Identity Clouds for secure access, authentication, and automation.

Okta describes its mission as enabling people to safely use technology across devices and applications.

Engineering opportunity

Okta is seeking an experienced Senior Site Reliability Engineer for the Emerging Products Group.

The team builds reliable, scalable, and secure cloud services using automation-first platform engineering and observability practices.

The role partners with software engineers, architects, and product teams to design, build, and operate cloud services.

Reliability and operations

  • Design, build, and operate large-scale cloud infrastructure and production services.
  • Participate in on-call support for highly available customer-facing systems.
  • Lead incident response and post-incident reviews.
  • Define and improve SLIs, SLOs, and error budgets.
  • Improve availability, scalability, performance, resilience, metrics, logging, tracing, dashboards, and alerting.

Engineering and automation

  • Develop software, automation, and infrastructure using Go, Python, Terraform, and related technologies.
  • Eliminate operational toil through automation, tooling, and platform engineering.
  • Improve deployment safety and workflows through CI/CD and GitOps.
  • Build self-service platforms, operational guardrails, and automation for developer velocity, reliability, and security.

Technical leadership and innovation

  • Drive reliability initiatives and guide engineers in operational best practices.
  • Mentor engineers through collaboration, design reviews, incident analysis, and knowledge sharing.
  • Support architecture and operational decisions with data-driven recommendations.
  • Explore AI-assisted engineering and emerging technologies to improve efficiency, incident response, troubleshooting, and automation.

Technology stack

  • Infrastructure and orchestration technologies include Kubernetes, EKS, GKE, Terraform, Helm, Git, ArgoCD, and GitOps.
  • Programming technologies include Golang and Python.
  • Observability technologies include Datadog and Splunk.
  • Data stores include PostgreSQL, Redis, and OpenSearch.

What we are looking for

  • Strong experience operating large-scale production services in AWS and/or GCP.
  • Deep expertise with Kubernetes in production environments.
  • Experience troubleshooting Kubernetes networking, storage, scheduling, scaling, and workload lifecycle issues.

Apply for this job in 1 click

Skip the repetitive application forms

Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.

Sarah M.James T.Maya R.

Trusted by over 500,000 job seekers on Base Career

Start Free Today

More from this employer

More jobs at Okta