Hiring.Camp

Senior Manager (DevOps, Automation, GCP)

CVS Health

·

2 weeks ago

Salary
$107k – $261k
Location
Work At Home-Connecticut, United States of America · Work At Home-North Carolina · Work At Home-Florida · Work At Home-Texas · Work At Home-Massachusetts
Workplace
Hybrid
Type
Full-time
Department
IT
Seniority
Senior
Experience
5+ years
Closing date
Today
Source
Workday

Description

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time.

Position Summary:
Join Fortune 7 CVS Health as a Sr. Manager, Software Engineering - DevOps, Observability & Monitoring to lead strategic initiatives for the CVS Caremark Digital team. In this role, you will drive the vision and execution of modern platform engineering capabilities that enable scalable, reliable, and secure application development. You will build and lead a high-performing engineering team responsible for designing, implementing, and operating cloud-native platforms. The ideal candidate is a hands-on leader with deep technical expertise in cloud architectures, DevOps, and observability, combined with the ability to lead transformation and operational excellence initiatives.
 

Key Responsibilities:

Leadership & People Management:

  • Lead, mentor, and grow a team of software engineers and SRE/DevOps engineers.

  • Foster a culture of accountability, innovation, and continuous improvement.

  • Define team goals, OKRs, and performance metrics aligned with organizational strategy.

  • Partner with product, architecture, and business stakeholders to deliver platform capabilities.

AIOps & Intelligent Operations:

  • Drive the adoption of AIOps solutions for predictive monitoring, anomaly detection, and automated root cause analysis.

  • Integrate machine learning models and analytics into monitoring pipelines to proactively detect and prevent incidents.

  • Develop intelligent alerting systems to reduce noise and improve signal quality.

Observability & Monitoring:

  • Architect and implement scalable observability frameworks across metrics, logs, traces, and events.

  • Establish standards for instrumentation, telemetry collection, and distributed tracing.

  • Enable proactive monitoring, alerting, and incident detection using tools such as Datadog, Prometheus, Grafana, and Splunk.

  • Define and implement enterprise-wide SRE practices, including SLIs, SLOs, error budgets, and reliability governance.

DevOps & Platform Engineering:

  • Drive adoption of CI/CD pipelines, Infrastructure as Code (IaC), and GitOps practices.

  • Lead the design and evolution of scalable, automated, and secure platform engineering solutions.

  • Standardize development and deployment workflows across teams.

  • Champion DevOps maturity, developer productivity, and release automation.

Reliability & Incident Management:

  • Improve system reliability through error budgets, resiliency patterns, and chaos engineering.

  • Lead incident response processes, postmortems, and root cause analysis.

  • Drive continuous improvements in MTTR (Mean Time to Recovery) and system availability.

 

Cloud & Infrastructure:

  • Architect and manage solutions across cloud platforms (AWS, Azure, or GCP).

  • Ensure scalability, security, and cost optimization of infrastructure.

  • Oversee containerization and orchestration using Docker and Kubernetes.

Security & Compliance:

  • Integrate DevSecOps practices into pipelines and platform tooling.

  • Ensure compliance with enterprise security standards and regulatory requirements.

  • Automate security scanning, vulnerability management, and policy enforcement.

Required Qualifications:

  • 5+ years of experience in software engineering, SRE, or production engineering within large-scale distributed systems.

  • Hands-on experience with AIOps or intelligent monitoring platforms, including anomaly detection and event correlation.

  • Experience with observability tools such as AppDynamics, Grafana, Prometheus, and Splunk.

  • Strong expertise in cloud platforms (AWS, Azure, or GCP), cloud-native architectures (Kubernetes, containers, microservices), and CI/CD pipelines (GitHub Actions, Jenkins).

  • Experience with Infrastructure as Code (Terraform, AWS CloudFormation, or GCP Deployment Manager).

  • Proficiency in at least one programming language (e.g., Python, Java, Go).

  • Proven track record of improving operational metrics (SLAs, CSAT, resolution time).

  • Experience with GenAI and automation tools such as OpenAI, Copilot, Gemini, Claude, and MCP.

  • Strong understanding of distributed systems, resiliency patterns, and fault tolerance.

  • Strong analytical skills with experience in reporting and performance measurement.

  • Excellent communication, stakeholder management, and conflict resolution skills.

  • Ability to thrive in a fast-paced, high-growth, or matrixed environment.

Preferred Qualifications:

  • Experience designing and implementing AIOps platforms or predictive reliability systems at scale.

  • Strong knowledge of machine learning applications in IT operations (e.g., anomaly detection, forecasting, clustering).

  • Experience defining and managing SLIs/SLOs and error budgets at scale.

  • Experience with OpenTelemetry and modern observability standards.

  • Familiarity with chaos engineering, resilience testing, and fault injection frameworks.

  • Exposure to GenAI-driven operations or AI-assisted troubleshooting tools.

  • Experience in healthcare, financial services, enterprise SaaS, or other regulated industries.

  • Proven ability to lead cross-functional initiatives and influence senior stakeholders.

  • Contributions to open-source projects related to SRE, observability, or AIOps.

  • Relevant certifications in cloud platforms, SRE, AIOps, OpenTelemetry, or DevOps are a plus.

Education:

  • Bachelor’s degree (or equivalent experience) in Computer Science, Engineering, or a related discipline.

 

Leadership Competencies:

  • Customer-first mindset with a passion for delivering exceptional service.

  • Strategic thinker with strong execution capabilities.

  • High emotional intelligence with strong people leadership skills.

  • Continuous improvement mindset with a focus on innovation.

Pay Range

The typical pay range for this role is:

$106,605.00 - $260,590.00


This pay range represents the base hourly rate or base annual full-time salary for all positions in the job grade within which this position falls.  The actual base salary offer will depend on a variety of factors including experience, education, geography and other relevant factors.  This position is eligible for a CVS Health bonus, commission or short-term incentive program in addition to the base pay range listed above.  This position also includes an award target in the company’s equity award program. 
 

Our people fuel our future. Our teams reflect the customers, patients, members and communities we serve and we are committed to fostering a workplace where every colleague feels valued and that they belong.

Great benefits for great people

We take pride in offering a comprehensive and competitive mix of pay and benefits that reflects our commitment to our colleagues and their families.

This full‑time position is eligible for a comprehensive benefits package designed to support the physical, emotional, and financial well‑being of colleagues and their families. The benefits for this position include medical, dental, and vision coverage, paid time off, retirement savings options, wellness programs, and other resources, based on eligibility.


Additional details about available benefits are provided during the application process and on
Benefits Moments.

We anticipate the application window for this opening will close on: 09/30/2026

Qualified applicants with arrest or conviction records will be considered for employment in accordance with all federal, state and local laws.

Skills

PythonJavaAWSAzureGCPDockerKubernetesTerraformJenkinsCI/CDMachine LearningGitHubSplunkDevOpsSREMicroservicesComplianceGo

Similar Jobs

30

Senior Manager, DevOps

Nttlimited · Johannesburg, South Africa · Remote

Today

Sr. Manager, DevOps

Accuweather · State College, PA +1

3 weeks ago

Senior DevOps Manager

O9Solutions · Dallas Office, United States of America

2 months ago

Senior Manager - SRE

LSEG · IND-BLR-Divyasree Technopolis, India

8 months ago

Senior Manager - DevOps

Quess · Noida, Uttar Pradesh, India

1+ year ago

Senior Manager Site Reliability Engineering

Akamai · United States, US

6 days ago

Senior Manager- Platform Engineering/DevOps

CVS Health · New York-161 Ave of the Americas, United States of America · Remote

1 week ago

Senior Engineering Manager - SRE

Oneadvanced · Bengaluru, KA, IN

1 week ago

Senior Engineering Manager - SRE

One Advanced · Bengaluru, KA, IN

1 week ago

Senior Technical Account Manager (DevOps)

Perforce Software · Minneapolis, MN · Hybrid

1 week ago

Senior Manager, Forward Deployed Engineer, Platform Engineering.

Jj · IN022 Hyderabad, India · Hybrid

1 week ago

Senior Manager, AI Engineer (Gen AI Platform Services: Agentic AI, Guardrails, Evaluation)

Capitalone · New York, NY, United States of America +4

2 weeks ago

Senior Vice President, Site Reliability Engineer Manager

0101022-GIA PROD US LOS ANGELES · Pune, MH, India

2 weeks ago

SRE - Enterprise & Cloud Security - AI Driven Security - Senior Manager

Pwc · New York - 300 Madison Avenue, United States of America +8

3 weeks ago

Senior Manager, Site Reliability Engineering (SRE)

Nium · Bangalore +1 · Hybrid

3 weeks ago

Tax Innovation - DevOps - Senior Manager

Pwc · Dallas - 2121 North Pearl Street, United States of America +3

3 weeks ago

Senior Middleware Platform Engineer - Data Transport - Manager

Statestreet · Bangalore, India · Hybrid

3 weeks ago

Senior Manager of Data Science Production Engineering, DevOps

Natera · San Carlos, CA

3 weeks ago

Senior Manager of Data Science Production Engineering, DevOps

Natera · US Remote · Remote

3 weeks ago

Sr. Manager – DevOps/QE

Rbc · MEADOWVALE BUSINESS PARK, 6880 FINANCIAL DR:MISSISSAUGA, Canada

3 weeks ago

Senior Manager, Site Reliability Engineering

Oracle · Ireland, IE

1 month ago

Senior Manager, Site Reliability Engineering – Paylo Platform

PDI Technologies · Alpharetta, GA +3 · Hybrid

1 month ago

Senior Manager, Site Reliability & Operational Resilience

Zelis Careers · US NJ Morristown, United States of America +4

1 month ago

Senior Manager, Machine Learning Platform Engineer

Gilead Sciences · US - CA - Foster City, United States of America · Onsite

1 month ago

Senior Manager, Machine Learning Platform Engineer

Gilead · US - CA - Foster City, United States of America · Onsite

1 month ago

Senior Manager, Site Reliability & Infrastructure Engineering

Aviva the most attractive choice · Canada - Markham ON 10 Aviva Way · Hybrid

1 month ago

Senior Manager, Site Reliability Engineering

Finastra · Mississauga - Avebury, Canada

1 month ago

Senior Manager of SRE

JPMorgan Chase · GLASGOW, LANARKSHIRE, United Kingdom, GB

1 month ago

Senior Manager of SRE

JP Morgan Chase · GLASGOW, LANARKSHIRE, United Kingdom, GB

1 month ago

Senior Manager, Site Reliability Engineering

Oracle · Reston, VA, United States, US

1 month ago