Manager, Site Reliability Engineering @Okta
All Others
Salary $182,000 — $250..
Remote Location
🇺🇸 USA Only
Employment Type full-time
Posted 1mth ago

[Hiring] Manager, Site Reliability Engineering @Okta

1mth ago - Okta is hiring a remote Manager, Site Reliability Engineering. 💸 Salary: $182,000 — $250,800 usd 📍Location: USA

Role Description

As a Manager, Site Reliability Engineer, you'll lead the SRE team with a focus on scalability, resilience, and empowering engineers to grow as technical leaders.

  • Lead the SRE team's technical direction, translating organizational vision into actionable roadmaps while driving complex, cross-functional initiatives across product and platform teams.
  • Operate at scale through hands-on participation in 24/7 on-call rotations (follow-the-sun weekdays, shared weekends), directly troubleshooting and remediating incidents on critical systems.
  • Build infrastructure resilience, designing and implementing monitoring, alerting, and automation improvements that reduce toil and elevate operational efficiency.
  • Champion reliability best practices, establishing policies and cultural standards that embed observability, resilience, and software engineering rigor into all engineering efforts.
  • Mentor and develop SRE talent, elevating team capabilities through pair programming, design discussions, and code reviews while fostering a culture of continuous learning.
  • Represent reliability as a senior technical leader in architectural reviews and strategic planning, ensuring reliability is a core consideration in major engineering decisions.

Qualifications

  • 3+ years of hands-on team leadership in SRE or software engineering roles within cloud-native environments, combined with 8+ years of total industry experience.
  • Deep expertise in cloud platforms (AWS, Azure) and infrastructure as code (Terraform), with proven experience managing cloud-native architectures including containers, Kubernetes, microservices, and databases.
  • Strong programming skills in Go or Python, with a track record of building and maintaining production-grade tools, automation, and infrastructure solutions.
  • Data-driven mindset grounded in SRE principles: blameless culture, systematic problem-solving, and the ability to apply software engineering approaches to operational challenges.
  • Exceptional communication skills—both verbal and written—enabling you to drive clarity during high-pressure incidents and articulate complex concepts to diverse stakeholders.
  • Proven ability to build and lead high-performing teams in globally distributed, remote-first environments with strong interpersonal and collaboration skills.
  • Strategic vision and technical depth, combining leadership acumen with hands-on technical excellence and a passion for mentoring senior engineers and shaping team direction.

Requirements

  • Experience leading reliability initiatives that directly improved system uptime and reduced incident response times at scale.
  • Contributions to open-source infrastructure or observability tooling.
  • Experience designing and implementing comprehensive incident response programs and runbook automation.
  • This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.

Benefits

  • Annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York, and Washington is between $182,000 — $250,800 USD.
  • Equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies.
Before You Apply
️
🇺🇸 Be aware of the location restriction for this remote position: USA Only
‼ Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Manager, Site Reliability Engineering @Okta
All Others
Salary $182,000 — $250..
Remote Location
🇺🇸 USA Only
Employment Type full-time
Posted 1mth ago
Apply for this position
Did not apply ✓
Applied ✓
Sent Follow-Up ✓
Interview Scheduled ✓
Interview Completed ✓
Offer Accepted ✓
Offer Declined ✓
Application Denied ✓
Unlock 125,000+ Remote Jobs
️
🇺🇸 Be aware of the location restriction for this remote position: USA Only
‼ Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply ✓
Applied ✓
Sent Follow-Up ✓
Interview Scheduled ✓
Interview Completed ✓
Offer Accepted ✓
Offer Declined ✓
Application Denied ✓
Unlock 125,000+ Remote Jobs
×
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 ★★★★★ from 500+ reviews

⚡ 127,048+ remote jobs, refreshed hourly

🔔 Real-time alerts: Apply first, direct to employer

🛡️ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later