- Location
- Bangalore, India
- Type
- Full-time
- Department
- IT
- Seniority
- Manager
- Source
- Workday
Description
We’re at the forefront of a once in a generational change in the broadband industry. Join us as we innovate, help our customers reach their potential, and connect underserved communities with unrivaled digital experiences.
As a Technical Program Manager in Cloud Operations, you will drive operational reliability across our Google Cloud Platform (GCP) ecosystem. You will lead critical incident post-mortems, govern security frameworks, and partner with application development and platform engineering teams to ensure all new services meet strict production entry criteria for reliability and observability.
Key Responsibilities:
Production Readiness & Operations Governance
- Define, govern, and enforce production entry criteria for all new services and features.
- Partner with application development and platform engineering to audit deployment readiness.
- Validate that incoming services meet corporate standards for observability, monitoring, and alerting.
- Scale reliable operational practices across engineering teams to reduce launch-day risks.
Critical Incident & Post-Mortem Management
- Own the end-to-end post-mortem lifecycle for critical GCP production outages and high-severity incidents.
- Lead "Five Whys" and blameless Root Cause Analysis (RCA) sessions to identify system vulnerabilities.
- Track actionable remediation items, ensuring timely execution to prevent recurrence.
- Report SLA/SLO compliance metrics and incident trends to engineering leadership.
Release Management & Compliance Coordination
- Bridge the gap between engineering, QA, and cloud operations to ensure streamlined CI/CD pipelines.
- Align release schedules with maintenance windows to minimize customer impact and system downtime.
- Partner with InfoSec to enforce Google Cloud security, data governance, and regulatory compliance.
- Maintain documentation and evidence collection for ongoing infrastructure compliance audits.
Qualifications:
- Experience: 5+ years of Technical Program Management, Systems Engineering, or Cloud Operations experience.
- GCP Expertise: Deep knowledge of Google Cloud services, security controls, and infrastructure.
- Production Readiness: Proven experience defining and enforcing production entry, reliability, and observability criteria.
- Incident Management: Strong track record running blameless post-mortems and driving infrastructure solutions from RCAs.
- Release & DevSecOps: Deep understanding of modern CI/CD tools, release management, and cloud-native architectures.
- Communication: Ability to align cross-functional engineering teams around shared operational and reliability goals.
Location:
- India – (Flexible hybrid work model - work from Bangalore office for 20 days in a quarter)