← Back to results

Site Reliability Engineer

Join Okta as a Site Reliability Engineer II to build reliable cloud services with a focus on automation and operational excellence.

Location
Bengaluru, India
Compensation
Not disclosed
Level
mid
Type
full time · Hybrid

Posted by employer 9 hours ago

First seen on Joblaze 4 hours ago

Last verified on the company career page 4 hours ago

What you'll build

  • Assist in operating large-scale cloud infrastructure
  • Participate in an SRE on-call rotation
  • Contribute to incident response efforts
  • Monitor and help maintain Service Level Indicators
  • Develop software, tools, and infrastructure-as-code updates

Must have

  • Hands-on experience operating production services in AWS and/or GCP environments
  • Solid, practical understanding of Kubernetes
  • Experience using Terraform and Helm
  • Proficiency in software development using Golang and/or Python
  • Good understanding of cloud networking fundamentals

Nice to have

  • Experience working with SaaS platforms
  • Experience deploying microservices in containerized environments
  • Familiarity with GitOps workflows and ArgoCD
  • Interest in exploring AI-assisted workflows

Practical constraints

  • Participate in an SRE on-call rotation

AI in the day-to-day

null

Not disclosed in this posting: compensation, years of experience, visa sponsorship.

Joblaze summary

In the role of Site Reliability Engineer II at Okta, the individual will focus on maintaining and enhancing the reliability of large-scale cloud infrastructure and production services. Key skills include proficiency in Go and Python, along with experience in Kubernetes, Terraform, and observability tools like Datadog. This position is suited for someone with a solid background in cloud operations and a strong desire to automate processes and improve system stability. The team emphasizes collaboration and continuous learning, making it ideal for engineers eager to grow their technical expertise.

Joblaze insights

  • Listed today — first seen on Joblaze October 6, 2026. Last confirmed on Okta's careers page October 6, 2026.

Quick facts

Is the Site Reliability Engineer role remote?
It's hybrid — Okta expects some on-site time in Bengaluru, India.
Where is the role based?
Okta is hiring for this position in Bengaluru, India.
What's the tech stack?
Joblaze extracted these technologies from the posting: Datadog, Golang, Kubernetes, OpenSearch, PostgreSQL, Python.
What seniority level is this role?
Okta targets mid-level candidates for this position.
Is this full-time or contract?
Full-time for this Site Reliability Engineer role at Okta.

From the original posting

Secure Every Identity, from AI to Human

Get to Know Okta

Okta is The World’s Identity Company. We free everyone to safely use any technology—anywhere, on any device or app. Our Workforce and Customer Identity Clouds enable secure yet flexible access, authentication, and automation that transforms how people move through the digital world, putting Identity at the heart of business security and growth.

At Okta, we celebrate a variety of perspectives and experiences. We are not looking for someone who checks every single box, we’re looking for lifelong learners and people who can make us better with their unique experiences.

Join our team! We’re building a world where Identity belongs to you.

The Engineering Opportunity

We are looking for a Site Reliability Engineer II to join Okta’s Emerging Products Group (EPG). Our mission is to build highly reliable, scalable, and secure cloud services that our customers can trust. We embrace an automation-first mindset and continuously invest in platform engineering, observability, and operational excellence to enable our engineering teams to move quickly and safely.

This role is ideal for a Site Reliability Engineer who is passionate about solving technical challenges, building automation, and maintaining the reliability of production systems. You will serve as a key technical contributor within the EPG SRE organization, collaborating closely with software engineers and senior SREs to support and improve world-class cloud services.

The ideal candidate aligns with the philosophy of “if you have to do it more than once, automate it” and possesses a strong drive for continuous learning, operational stability, and software engineering.

What You’ll Be Doing

Reliability & Operations

  • Support and Maintain: Assist in operating large-scale cloud infrastructure and production services, focusing on deployment, maintenance, and day-to-day health.
  • On-Call Rotation: Participate in an SRE on-call rotation supporting highly available customer-facing systems.
  • Incident Response: Contribute to incident response efforts, assist in troubleshooting, and participate in post-incident reviews to identify systemic fixes.
  • Service Metrics: Monitor and help maintain Service Level Indicators (SLIs) and Service Level Objectives (SLOs), raising awareness when error budgets are threatened.
  • Observability: Implement and configure observability configurations using metrics, logging, tracing, dashboards, and alerting patterns established by the team.

Engineering & Automation

  • Automation Development: Develop software, tools, and infrastructure-as-code updates using Go, Python, Terraform, and related technologies.
  • Toil Reduction: Identify opportunities to eliminate operational toil through automation and self-service tooling.
  • CI/CD & GitOps: Utilize and maintain CI/CD pipelines and GitOps workflows to ensure safe, repeatable deployment of infrastructure and services.
  • Platform Alignment: Collaborate with engineering teams to migrate and integrate existing workloads into modern platform capabilities.

Collaboration & Execution

  • Project Execution: Execute assigned technical tasks and sub-projects from design through to production rollout with guidance from senior SREs.
  • Knowledge Sharing: Document operational procedures, runbooks, and share technical knowledge within the team.
  • Standard Adherence: Actively follow and adopt operational best practices, security standards, and reliability principles.

Innovation

  • Modern Tooling: Learn and apply modern operational techniques—including AI-assisted tools—to increase daily productivity, accelerate troubleshooting, and simplify automation workflows.

Our Tech Stack

Category

Technologies

Infrastructure/Orchestration

Kubernetes (EKS/GKE), Terraform, Helm, Git, ArgoCD, GitOps

Programming

Golang, Python

Observability

Datadog, Splunk

Data Stores

PostgreSQL, Redis, OpenSearch

What We Are Looking For

Technical Excellence

  • Cloud Operations: Hands-on experience operating production services in AWS and/or GCP environments.
  • Kubernetes Foundation: Solid, practical understanding of Kubernetes, container runtimes, and troubleshooting application workloads.
  • Infrastructure as Code: Experience using Terraform and Helm to provision and manage cloud resources.
  • Programming Skills: Proficiency in software development using Golang and/or Python.
  • Networking Basics: Good understanding of cloud networking fundamentals, including DNS, load balancing, HTTPS/TLS, and basic traffic routing.
  • Data Platforms: Familiarity operating and querying distributed data stores (e.g., PostgreSQL, Redis, OpenSearch).
  • Observability: Hands-on experience configuring metrics, alerts, and dashboards on platforms like Datadog or Splunk.

Operational Excellence

  • Production Experience: Experience supporting customer-facing production systems.
  • Problem Solving: Proven ability to logically troubleshoot complex infrastructure and application issues.
  • Reliability Concepts: Understanding of fundamental SRE concepts, such as SLOs, SLA boundaries, and blameless post-mortem cultures.
  • Continuous Integration: Familiarity with modern CI/CD pipelines and automated deployment strategies.

Security & Compliance

  • Cloud Security: Basic understanding of cloud security fundamentals, including Identity & Access Management (IAM) and secure secrets management.

Professional & Team Skills

  • Strong Collaboration: Excellent communication and collaboration skills to work effectively with globally distributed engineering teams.
  • Growth Mindset: High motivation to learn from senior engineers, accept technical feedback, and continually grow technical capabilities.
  • Execution Focus: Proven ability to deliver assigned tasks on time while maintaining work quality.

Preferred Qualifications

  • Experience working with SaaS platforms.
  • Experience deploying microservices in containerized environments.
  • Familiarity with GitOps workflows and ArgoCD.
  • Interest in exploring AI-assisted workflows or modern developer productivity tools.

#LI-Hybrid
#P25718

The Okta Experience

Standard company text repeated across Okta's postings is omitted here.

Similar positions

Okta
Staff Site Reliability Engineer
Okta · Bengaluru, India
Okta
Senior Site Reliability Engineer
Okta · Bengaluru, India
Okta
Principal Site Reliability Engineer
Okta · Bengaluru, India
Okta
Senior Site Reliability Engineer (FedRAMP)
Okta · San Francisco, California