Director, Site Reliability Engineering
2 days ago
Manchester
\n Why Join Omnicell?\n At Omnicell, we’re transforming the future of medication management through intelligent automation, cloud‑native platforms, and innovative healthcare technology. Every service we build and every platform we operate ultimately contributes to improving patient safety, reducing clinician burden, and enabling better healthcare outcomes worldwide. About This Opportunity\n This is not a traditional operations leadership position. \n This is an opportunity to build the engineering organization that defines how Omnicell designs, measures, automates, and continuously improves reliability across its cloud platforms. \n Reporting directly to the Vice President of Global Cloud Operations, the Director of Site Reliability Engineering will establish Omnicell’s enterprise reliability engineering strategy, build a globally distributed SRE organization, and partner closely with Cloud Platform Engineering, Site Reliability Operations, Cloud Security, Product Engineering, and Enterprise Architecture to ensure reliability is engineered into every service throughout its lifecycle. \n Working alongside the Director of Site Reliability Operations, this leader will define the engineering standards, automation, observability, production readiness, and resilience capabilities that enable world‑class operational performance. While Site Reliability Operations owns the day‑to‑day operation of production services, the Site Reliability Engineering organization is responsible for continuously improving the architecture, automation, and engineering practices that make those services resilient, scalable, and highly available. \n The successful candidate will be equally comfortable discussing executive strategy with senior leadership, reviewing Kubernetes architectures with engineering teams, mentoring engineering leaders, and influencing enterprise technology direction. Purpose\n The Director of Site Reliability Engineering is responsible for establishing and leading Omnicell’s global Site Reliability Engineering organization. \n This role owns the engineering strategy, governance, architecture, and technical practices that improve service reliability through software engineering, observability, automation, resilience engineering, production engineering, and operational readiness. \n Rather than operating production systems on a day‑to‑day basis, the Site Reliability Engineering organization develops the platforms, standards, tooling, automation, and engineering capabilities that enable highly reliable cloud services across the enterprise. \n Success in this role is measured by how effectively reliability is designed, engineered, automated, and continuously improved across Omnicell’s cloud ecosystem. Primary Impact\n As Director of Site Reliability Engineering, you will define how reliability is engineered across Omnicell’s cloud platform. \n Your leadership will enable engineering teams to design, build, deploy, and continuously improve resilient cloud services by establishing enterprise reliability standards, automation frameworks, observability strategies, production engineering practices, and engineering governance. \n Working closely with Site Reliability Operations, your organization will reduce operational complexity by engineering solutions that eliminate manual effort, improve resilience, automate recovery, and continuously improve platform reliability. \n By building a culture centered on engineering excellence, automation, continuous learning, and measurable operational outcomes, you will help create a cloud platform capable of supporting Omnicell’s long‑term product strategy and business growth. What You’ll DoBuild and Lead a World-Class Site Reliability Engineering Organization\n Build and scale Omnicell’s global Site Reliability Engineering organization. You Will\n \n • Recruit, mentor, and develop SRE Managers, Principal Engineers, and senior technical talent.\n, • Build a high‑performing engineering organization focused on reliability, resilience, and production engineering.\n, • Define engineering standards, career development frameworks, and technical leadership expectations.\n, • Foster a culture of ownership, accountability, continuous improvement, and engineering excellence.\n, • Develop service‑aligned SRE engagement models that partner closely with Product Engineering teams.\n Define Enterprise Reliability Engineering Strategy\n Develop Omnicell’s enterprise reliability engineering strategy by establishing standards and governance for: \n \n • Site Reliability Engineering\n, • Production Engineering\n, • Reliability Engineering\n, • Resilience Engineering\n, • Production Architecture\n, • Operational Readiness Engineering\n, • Availability Engineering\n, • Capacity Engineering\n, • Service Reliability Reviews\n \n Develop multi‑year engineering roadmaps that improve service reliability while enabling engineering teams to deliver software with greater speed and confidence. Engineering ReliabilityOwn Omnicell’s Enterprise Reliability Framework, Including\n \n • Service Level Indicators (SLIs)\n, • Service Level Objectives (SLOs)\n, • Error Budgets\n, • Reliability Design Standards\n, • Production Readiness Reviews\n, • Capacity Planning Models\n, • Failure Mode Analysis\n, • Reliability Scorecards\n, • Engineering Guardrails\n \n Partner with Product Engineering and Cloud Platform Engineering to embed reliability throughout the software development lifecycle. Production Engineering\n Establish a Production Engineering practice focused on continuously improving the operational characteristics of Omnicell’s cloud services. Responsibilities Include\n \n • Performance Engineering\n, • Scalability Engineering\n, • Availability Engineering\n, • Capacity Planning\n, • Service Hardening\n, • Resilience Testing\n, • Operational Readiness Reviews\n, • Production Design Reviews\n, • Failure Analysis\n, • Continuous Reliability Improvement\n \n Partner with Product Engineering throughout the software development lifecycle to ensure every service meets enterprise production standards before deployment. Observability Engineering\n Establish Omnicell’s enterprise observability engineering strategy. Own Engineering Standards Supporting\n \n • OpenTelemetry\n, • Metrics Architecture\n, • Distributed Tracing\n, • Centralized Logging\n, • Application Performance Monitoring\n, • Synthetic Monitoring\n, • Alert Engineering\n, • Dashboard Standards\n, • Service Health Models\n, • Telemetry Architecture\n \n Partner with Cloud Platform Engineering and Site Reliability Operations to ensure telemetry provides actionable operational intelligence while enabling proactive reliability improvements. Reliability Automation\n Champion an engineering‑first approach to automation. Lead Initiatives That\n \n • Eliminate operational toil through software engineering.\n, • Develop self‑healing platform capabilities.\n, • Build automated remediation workflows.\n, • Improve deployment safety.\n, • Increase service resilience.\n, • Enhance engineering productivity.\n, • Expand predictive reliability capabilities.\n, • Enable AI‑assisted engineering workflows.\n \n Partner with Cloud Platform Engineering to integrate automation into the Internal Developer Platform while supporting Site Reliability Operations through engineering‑driven operational automation. Engineering PartnershipDevelop Trusted Partnerships Across\n \n • Product Engineering\n, • Cloud Platform Engineering\n, • Site Reliability Operations\n, • Cloud Security\n, • Enterprise Architecture\n, • Quality Engineering\n, • Technical Support\n \n Provide engineering leadership that enables operational excellence while ensuring reliability remains a shared responsibility across the software delivery lifecycle. Organizational LeadershipProvide Strategic Leadership For The SRE Organization By\n \n • Defining organizational strategy and operating model.\n, • Establishing enterprise engineering standards.\n, • Developing future engineering leaders.\n, • Driving employee engagement and career development.\n, • Building an inclusive, high‑performing engineering culture.\n, • Promoting innovation and continuous learning.\n Executive LeadershipPartner With Executive Leadership To\n \n, • Present enterprise reliability metrics and engineering KPIs.\n, • Recommend strategic technology investments.\n, • Communicate platform reliability trends and technical risks.\n, • Support enterprise cloud transformation initiatives.\n, • Influence engineering strategy across Product, Platform, Security, and Operations.\n, • Represent Site Reliability Engineering during executive planning and technology reviews.\n What Success Looks LikeFirst 90 Days\n \n, • Build relationships with Engineering, Product, Security, Platform Engineering, and Site Reliability Operations leaders.\n, • Assess current reliability engineering maturity.\n, • Evaluate production architecture, observability, and automation capabilities.\n, • Identify engineering opportunities to improve platform resilience.\n, • Develop a three‑year Site Reliability Engineering roadmap.\n First Six Months\n \n, • Establish enterprise SLO, SLI, and Error Budget framework.\n, • Launch Production Engineering practice.\n, • Standardize Production Readiness Reviews.\n, • Define enterprise observability engineering standards.\n, • Recruit key Site Reliability Engineering leaders.\n, • Publish enterprise reliability scorecards.\n First Year\n \n, • Build a high‑performing global Site Reliability Engineering organization.\n, • Establish reliability engineering as a core engineering discipline.\n, • Demonstrate measurable improvements in platform availability, deployment safety, and engineering productivity.\n, • Reduce operational toil through engineering automation.\n, • Deploy enterprise observability standards across all strategic platforms.\n, • Establish a scalable engineering operating model supporting Omnicell’s long‑term cloud strategy.\n Leadership ExpectationsAs Director Of Site Reliability Engineering, You Will\n \n, • Establish the technical vision for enterprise reliability engineering.\n, • Build and mentor a world‑class Site Reliability Engineering organization.\n, • Drive engineering excellence through architecture, automation, observability, and production engineering.\n, • Influence enterprise engineering strategy across Product, Platform, Security, and Operations.\n, • Foster a culture of continuous learning, innovation, and measurable improvement.\n, • Serve as Omnicell’s executive authority on reliability engineering, production architecture, and cloud resilience.\n \n #J-18808-Ljbffr