okta

Manager, Site Reliability Engineering

Apply Now

At a Glance

Location
San Francisco, California, United States
Work Regime
hybrid
Experience
3+ years
Posted
2026-07-02T09:09:41-04:00

Key Requirements

Required Skills

AWSAgileCI/CDDevOpsKubernetesTerraform

Domain Knowledge

  • Automation
  • Insurance
  • SaaS

Requirements

Extensive experience using Agile and DevOps methodologies to build product infrastructure and shared service at scale

Experience running large-scale infrastructure platforms supporting a SaaS/Cloud service in a public Cloud, preferably AWS.

Experience supporting a multi-Cloud environment will be a plus.

Strong expertise in cloud-native architectures, containerization (Kubernetes), IaC (Terraform), and CI/CD pipelines

Strong background and hands-on experience in SW development, PaaS and automation

Deep experience with building and operating observability platforms and monitoring tools (Grafana, Splunk, APM etc.) in a large scale environment.

Compensation & Benefits

$204,000

$306,000 USD

The Okta Experience

Supporting Your Well-Being

Driving Social Impact

Responsibilities

Managing a team of SRE’s supporting various workloads and teams that support our IDaaS platform.

Drive the microservice journey, DevOps maturity, and workload reliability in tandem with architects and teams across the organization.

Accelerate the velocity of SRE and product engineering by developing powerful tooling, intuitive self-service capabilities, and robust self-healing patterns.

Lead, mentor, and grow a high-performing team of engineers and managers across platform, infrastructure, and shared services domains.

Perform engineering design evaluations and ensure the completion of projects within resource, budget, and scheduling constraints.

Improve SDLC processes for Cloud infrastructure as a code, including the maturity of CI/CD pipelines, change and release management