Jobs / CVS Health
Executive Director - Site Reliability Engineering - Retail Pharmacy
CVS Health · Woonsocket, RI, United States
Woonsocket, RI, United StatesExp: 15+ yrs175,100-334,750 USD/yearlyRemote
Remuneration
175,100-334,750 USD/yearly
Location
Woonsocket, RI, United States
Visa sponsorship
Not specified
Job summary
The Executive Director of Site Reliability Engineering at CVS Health is responsible for the reliability, resilience, and performance of the retail and pharmacy technology ecosystem. This role involves defining a comprehensive reliability strategy, overseeing global engineering teams, and championing modern SRE practices to enhance operational excellence.
Qualifications
- 15+ years of technology leadership experience in infrastructure, operations, or software engineering.
- Experience leading Site Reliability Engineering or large-scale reliability programs.
- Success managing complex, distributed technology environments.
- Expertise in incident management, observability, and operational excellence.
- Understanding of retail technology ecosystems including POS and pharmacy applications.
- Ability to lead enterprise-wide transformation initiatives.
- Exceptional communication and problem-solving skills.
- Experience leading large-scale global teams.
- Commitment to collaboration, inclusion, and innovation.
Responsibilities
- Define and lead the enterprise-wide Site Reliability Engineering strategy.
- Align reliability and operational objectives with business and technology priorities.
- Develop and execute a multi-year roadmap for observability, automation, and operational excellence.
- Influence technology investment decisions and architectural direction.
- Establish and govern Service Level Objectives (SLOs) and Service Level Agreements (SLAs).
- Drive continuous improvements in availability, performance, and resilience.
- Lead major incident management activities for rapid detection and remediation.
- Foster a culture of accountability and continuous improvement.
- Establish observability capabilities including monitoring and logging.
- Deliver visibility into the health and performance of systems.
- Drive automation initiatives to reduce manual effort.
- Leverage AI and machine learning for monitoring and incident prevention.
- Champion modern cloud and distributed systems architectures.
- Partner with teams to embed reliability practices in the software development lifecycle.
- Ensure technology operations are secure and compliant.
- Lead and mentor global teams of Site Reliability Engineers.
- Build an inclusive engineering culture focused on innovation.
- Develop workforce planning strategies to attract and retain talent.
- Promote leadership development and succession planning.
Skills
AWSAzureDatadogDynatraceGCPGrafanaKubernetesOpenShiftPrometheusSplunk
Degrees
Bachelor's degree in Computer ScienceEngineeringInformation Technology
Relocation
No