okta
Manager, Site Reliability Engineering
At a Glance
- Location
- San Francisco, California, United States
- Work Regime
- hybrid
- Experience
- 3+ years
- Posted
- 2026-07-02T09:09:41-04:00
Key Requirements
Required Skills
Domain Knowledge
- Automation
- Insurance
- SaaS
Requirements
Extensive experience using Agile and DevOps methodologies to build product infrastructure and shared service at scale
Experience running large-scale infrastructure platforms supporting a SaaS/Cloud service in a public Cloud, preferably AWS.
Experience supporting a multi-Cloud environment will be a plus.
Strong expertise in cloud-native architectures, containerization (Kubernetes), IaC (Terraform), and CI/CD pipelines
Strong background and hands-on experience in SW development, PaaS and automation
Deep experience with building and operating observability platforms and monitoring tools (Grafana, Splunk, APM etc.) in a large scale environment.
Compensation & Benefits
$204,000
—
$306,000 USD
The Okta Experience
Supporting Your Well-Being
Driving Social Impact
Responsibilities
Managing a team of SRE’s supporting various workloads and teams that support our IDaaS platform.
Drive the microservice journey, DevOps maturity, and workload reliability in tandem with architects and teams across the organization.
Accelerate the velocity of SRE and product engineering by developing powerful tooling, intuitive self-service capabilities, and robust self-healing patterns.
Lead, mentor, and grow a high-performing team of engineers and managers across platform, infrastructure, and shared services domains.
Perform engineering design evaluations and ensure the completion of projects within resource, budget, and scheduling constraints.
Improve SDLC processes for Cloud infrastructure as a code, including the maturity of CI/CD pipelines, change and release management