← Back to results

Staff SRE for K8s Platform Team (AWS, Kubernetes, Platform Creation, Helm, Karpenter, Istio)

Join Okta as a Staff SRE to architect and manage Kubernetes platforms on AWS, ensuring high availability and performance.

Location
Bengaluru, India
Compensation
Not disclosed
Level
staff
Type
full time · Hybrid

Posted by employer 2 months ago

First seen on Joblaze 2 months ago

Last verified on the company career page 1 hour ago

Requirements

Experience
8+ years
Education
Bachelor's degree

Not disclosed in this posting: compensation, visa sponsorship.

Joblaze summary

The Staff Site Reliability Engineer at Okta is responsible for designing and managing robust Kubernetes platforms that support cloud-native applications, ensuring they are reliable, scalable, and secure. Key skills include expertise in AWS infrastructure, Kubernetes management, and automation tools like Helm and Karpenter. This role is suited for seasoned professionals with extensive experience in cloud technologies and a strong background in infrastructure-as-code practices. Okta emphasizes a collaborative environment, aiming to innovate in identity security.

Joblaze insights

  • Listed about 2 months ago — first seen on Joblaze August 3, 2026. Last confirmed on Okta's careers page October 9, 2026.
  • Kubernetes appears in 68% of 150 comparable staff devops/sre roles; Karpenter appears in 0.7% of 150 comparable staff devops/sre roles.

Quick facts

Is the Staff SRE for K8s Platform Team (AWS, Kubernetes, Platform Creation, Helm, Karpenter, Istio) role remote?
It's hybrid — Okta expects some on-site time in Bengaluru, India.
How much experience is required?
At least 8 years of relevant experience for this Staff SRE for K8s Platform Team (AWS, Kubernetes, Platform Creation, Helm, Karpenter, Istio) role.
Where is the role based?
Okta is hiring for this position in Bengaluru, India.
What's the tech stack?
Joblaze extracted these technologies from the posting: API Gateway, AWS, AWS Lambda, Ansible, CI/CD, Chef.
What seniority level is this role?
Okta targets staff-level candidates for this position.
Is this full-time or contract?
Full-time for this Staff SRE for K8s Platform Team (AWS, Kubernetes, Platform Creation, Helm, Karpenter, Istio) role at Okta.

From the original posting

Secure Every Identity, from AI to Human

Workforce Identity Cloud
Okta Workforce Identity Cloud (WIC) provides easy, secure access for your workforce so you can focus on other strategic priorities, such as reducing costs and doing more for your customers.
If you like to be challenged and have a passion for solving large-scale automation, testing, and tuning problems, we would love to hear from you. The ideal candidate is someone who exemplifies the ethics of, “If you have to do something more than once, automate it” and who can rapidly self-educate on new concepts and tools.
Position Overview:
The Staff Site Reliability Engineer (SRE) will play a key role in building and managing Kubernetes platforms that support cloud-native applications and services. This position focuses on architecting and managing reliable, scalable, and secure Kubernetes-based platforms on AWS, ensuring high availability and performance while optimising costs and automation. The ideal candidate will have hands-on experience with AWS infrastructure, Kubernetes platform creation, Helm charts, Karpenter scaling, and Istio service mesh.
Key Responsibilities:
  • Kubernetes Platform Creation: Design, implement, and maintain highly available, scalable, and fault-tolerant Kubernetes platforms. Ensure clusters are optimised for production workloads, providing high resilience and operational efficiency.
  • AWS Infrastructure Management: Build, manage, and optimise AWS cloud infrastructure, including EKS, ECS, S3, VPCS, RDS, IAM, and more. Implement best practices for cost management, scaling, and security within AWS.
  • Helm Management: Utilise Helm to automate and streamline the deployment of applications and services to Kubernetes clusters. Create, maintain, and manage Helm charts for production-ready deployments.
  • Karpenter Implementation: Implement and manage Karpenter to dynamically scale Kubernetes clusters in response to workload demands.
  • Istio Service Mesh Management: Configure and manage Istio to provide service-to-service communication, security, and observability within the Kubernetes clusters. Enable fine-grained traffic management, service discovery, and policy enforcement.
  • Platform Automation & Scaling: Automate the deployment, scaling, and management of infrastructure and applications. Work with CI/CD pipelines to ensure a seamless flow from development to production with minimal downtime.
  • Incident Management & Troubleshooting: Respond to incidents, troubleshoot, and resolve system issues related to performance, availability, and security in a timely and effective manner.
  • Security & Compliance: Design and implement secure cloud infrastructure with appropriate access controls, network security, and compliance frameworks.
  • Documentation & Knowledge Sharing: Create and maintain detailed documentation for Kubernetes platform setup, operational procedures, and best practices. Promote knowledge sharing across teams.
Required Qualifications:
  • 5+ years of experience with Kubernetes/ K8s, Helm,Karpenter,Istio;
  • 8+ years of Experience with infrastructure-as-code tools like Terraform, Chef or Ansible
  • 8+ years of Experience with serverless computing (AWS Lambda, API Gateway) and microservices architecture.
  • Proven experience with AWS (EKS, ECS, RDS, S3, CloudFormation, IAM, etc.) and solid understanding of cloud-native architectures.
  • Strong expertise in Kubernetes platform creation, management, and optimisation (e.g., setting up highly available clusters, networking, and storage).
  • Hands-on experience with Helm for Kubernetes application deployment and management.
  • Practical experience with Karpenter for dynamic scaling of Kubernetes clusters and optimising resource usage.
  • Expertise in managing and securing Istio for service mesh, including traffic management, security, and observability features.
  • Proficiency in CI/CD pipelines and automation tools (e.g., Jenkins, GitLab, CircleCI, Terraform, Spinnaker, Ansible).
  • Strong scripting and automation skills in Python or Go for infrastructure management and platform automation.
  • Experience with monitoring, logging, and alerting tools such as Prometheus, Grafana, CloudWatch, and ELK Stack.
Preferred Qualifications:
  • Experience with multi-region cloud environments.
  • Understanding of security best practices for cloud platforms and Kubernetes (e.g., role-based access control (RBAC), encryption, and compliance frameworks).
  • Familiarity with Docker and containerization principles.
  • Bachelor’s degree in Computer Science, Engineering, or related field (or equivalent professional experience).
  • Certifications (Preferred): CKA (Certified Kubernetes Administrator), CKAD (Certified Kubernetes Application Developer), or AWS Certified DevOps Engineer are highly desirable.

#LI_Hybrid

P24488_3510130

The Okta Experience

Standard company text repeated across Okta's postings is omitted here.

Similar positions

Okta
Staff Site Reliability Engineer - Kubernetes
Okta · Bellevue, Washington; Chicago, Illinois; New York, New York; San Francisco, California; Washington, DC
Okta
Staff Site Reliability Engineer
Okta · Bengaluru, India
Okta
Staff Site Reliability Engineer
Okta · Bengaluru, India
Okta
Senior Site Reliability Engineer
Okta · Bengaluru, India