Director of Site Reliability Engineering @Omnicell
All Others
Salary unspecified
Remote Location
remote UK
Employment Type full-time
Posted 1mth ago

[Hiring] Director of Site Reliability Engineering @Omnicell

1mth ago - Omnicell is hiring a remote Director of Site Reliability Engineering. 💸 Salary: unspecified 📍Location: UK

Role Description

The Director of Site Reliability Engineering will have a unique opportunity to build and lead a world-class SRE organization responsible for ensuring our cloud platforms deliver exceptional availability, resilience, performance, and customer experience. This is a transformational leadership role for someone passionate about building engineering organizations, driving operational excellence through software engineering, and influencing technology strategy at an enterprise level.

What You’ll Do

  • Build and Lead a World-Class Site Reliability Engineering Organization
    • Recruit, mentor, and develop SRE Managers, Principal Engineers, and senior technical talent.
    • Build a high-performing engineering organization focused on reliability, resilience, and production engineering.
    • Define engineering standards, career development frameworks, and technical leadership expectations.
    • Foster a culture of ownership, accountability, continuous improvement, and engineering excellence.
    • Develop service-aligned SRE engagement models that partner closely with Product Engineering teams.
  • Define Enterprise Reliability Engineering Strategy
    • Develop Omnicell’s enterprise reliability engineering strategy by establishing standards and governance for:
      • Site Reliability Engineering
      • Production Engineering
      • Reliability Engineering
      • Resilience Engineering
      • Production Architecture
      • Operational Readiness Engineering
      • Availability Engineering
      • Capacity Engineering
      • Service Reliability Reviews
    • Develop multi-year engineering roadmaps that improve service reliability while enabling engineering teams to deliver software with greater speed and confidence.
  • Engineering Reliability
    • Own Omnicell’s enterprise reliability framework, including:
      • Service Level Indicators (SLIs)
      • Service Level Objectives (SLOs)
      • Error Budgets
      • Reliability Design Standards
      • Production Readiness Reviews
      • Capacity Planning Models
      • Failure Mode Analysis
      • Reliability Scorecards
      • Engineering Guardrails
    • Partner with Product Engineering and Cloud Platform Engineering to embed reliability throughout the software development lifecycle.
  • Production Engineering
    • Establish a Production Engineering practice focused on continuously improving the operational characteristics of Omnicell’s cloud services. Responsibilities include:
      • Performance Engineering
      • Scalability Engineering
      • Availability Engineering
      • Capacity Planning
      • Service Hardening
      • Resilience Testing
      • Operational Readiness Reviews
      • Production Design Reviews
      • Failure Analysis
    • Partner with Product Engineering throughout the software development lifecycle to ensure every service meets enterprise production standards before deployment.
  • Observability Engineering
    • Establish Omnicell’s enterprise observability engineering strategy. Own engineering standards supporting:
      • OpenTelemetry
      • Metrics Architecture
      • Distributed Tracing
      • Centralized Logging
      • Application Performance Monitoring
      • Synthetic Monitoring
      • Alert Engineering
      • Dashboard Standards
      • Service Health Models
      • Telemetry Architecture
    • Partner with Cloud Platform Engineering and Site Reliability Operations to ensure telemetry provides actionable operational intelligence while enabling proactive reliability improvements.
  • Reliability Automation
    • Champion an engineering-first approach to automation. Lead initiatives that:
      • Eliminate operational toil through software engineering.
      • Develop self-healing platform capabilities.
      • Build automated remediation workflows.
      • Improve deployment safety.
      • Increase service resilience.
      • Enhance engineering productivity.
      • Expand predictive reliability capabilities.
      • Enable AI-assisted engineering workflows.
    • Partner with Cloud Platform Engineering to integrate automation into the Internal Developer Platform while supporting Site Reliability Operations through engineering-driven operational automation.
  • Engineering Partnership
    • Develop trusted partnerships across:
      • Product Engineering
      • Cloud Platform Engineering
      • Site Reliability Operations
      • Cloud Security
      • Enterprise Architecture
      • Quality Engineering
      • Technical Support
    • Provide engineering leadership that enables operational excellence while ensuring reliability remains a shared responsibility across the software delivery lifecycle.
  • Organizational Leadership
    • Provide strategic leadership for the SRE organization by:
      • Defining organizational strategy and operating model.
      • Establishing enterprise engineering standards.
      • Developing future engineering leaders.
      • Driving employee engagement and career development.
      • Building an inclusive, high-performing engineering culture.
      • Promoting innovation and continuous learning.
  • Executive Leadership
    • Partner with executive leadership to:
      • Present enterprise reliability metrics and engineering KPIs.
      • Recommend strategic technology investments.
      • Communicate platform reliability trends and technical risks.
      • Support enterprise cloud transformation initiatives.
      • Influence engineering strategy across Product, Platform, Security, and Operations.
      • Represent Site Reliability Engineering during executive planning and technology reviews.

Qualifications

  • Experience in building and leading high-performing engineering teams.
  • Strong background in Site Reliability Engineering, Production Engineering, and Cloud Technologies.
  • Proven ability to define and implement engineering standards and best practices.
  • Excellent communication and leadership skills.
  • Experience with cloud platforms and modern software development practices.

Requirements

  • Experience with Kubernetes and cloud-native architectures.
  • Strong understanding of observability and reliability engineering principles.
  • Ability to mentor and develop engineering talent.
  • Experience in strategic planning and execution.
  • Proven track record of driving operational excellence.

Benefits

  • Competitive salary and performance-based bonuses.
  • Comprehensive health benefits.
  • Retirement savings plan with company match.
  • Flexible work arrangements.
  • Professional development opportunities.
Before You Apply
️
remote Be aware of the location restriction for this remote position: UK
‼ Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Director of Site Reliability Engineering @Omnicell
All Others
Salary unspecified
Remote Location
remote UK
Employment Type full-time
Posted 1mth ago
Apply for this position
Did not apply ✓
Applied ✓
Sent Follow-Up ✓
Interview Scheduled ✓
Interview Completed ✓
Offer Accepted ✓
Offer Declined ✓
Application Denied ✓
Unlock 125,000+ Remote Jobs
️
remote Be aware of the location restriction for this remote position: UK
‼ Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply ✓
Applied ✓
Sent Follow-Up ✓
Interview Scheduled ✓
Interview Completed ✓
Offer Accepted ✓
Offer Declined ✓
Application Denied ✓
Unlock 125,000+ Remote Jobs
×
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 ★★★★★ from 500+ reviews

⚡ 127,423+ remote jobs, refreshed hourly

🔔 Real-time alerts: Apply first, direct to employer

🛡️ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later