- Location
- Atlanta, GA · Atlanta, Georgia, United States
- Department
- 51850 Consumer QA & Dev Ops
- Seniority
- Senior
- Experience
- 8+ years
- Source
- Greenhouse
Description
About The Weather Company:
The Weather Company is the world’s leading weather provider, helping people and businesses make more informed decisions and take action in the face of weather. Together with advanced technology and AI, The Weather Company’s high-volume weather data, insights, advertising, and media solutions across the open web help people, businesses, and brands around the world prepare for and harness the power of weather in a scalable, privacy-forward way. The world’s most accurate forecaster globally, the company reaches hundreds of enterprise clients and more than 360 million monthly active users via its digital properties from The Weather Channel (weather.com) and Weather Underground (wunderground.com).
AI is part of how we work:
Human judgment, expertise, and creativity remain essential and are amplified by AI. We expect everyone in this role to use AI thoughtfully and proactively to accelerate their performance, improve the quality of their work, and explore new possibilities. We are looking for people who are curious, adaptable, and excited to help shape an AI-enabled culture that delivers better outcomes for our customers and our company.
Job brief:
We are seeking an experienced and vision-driven Senior Manager, DevOps Engineering to lead, mentor, and scale our global DevOps, SRE, and Infrastructure engineering teams. In this role, you will champion modern engineering culture, drive cloud modernization, and ensure our infrastructure is secure, highly available, and resilient.
You will partner closely with Software Engineering, Product, Security, and IT Operations to deliver zero-downtime releases, mature our CI/CD and IaC practices, and establish robust release and environment management frameworks.
The impact you'll make:
- Lead, mentor, and develop a high-performing global team of DevOps engineers, Site Reliability Engineers (SREs), and infrastructure engineers, to lead transformation and foster a culture of collaboration, accountability, and continuous improvement.
- Oversee the architecture, deployment, and operational management of cloud environments (AWS) to ensure high availability, scalability, security, and cost optimization.
- Define and enforce best practices for Infrastructure as Code (IaC), specifically with AWS experience, using tools like Terraform and CloudFormation, ensuring consistent and reliable provisioning of resources.
- Drive the adoption and optimization of CI/CD pipelines, containerization technologies (Docker, Kubernetes), and orchestration platforms to streamline software delivery processes.
- Collaborate closely with software engineering, product, security, and IT operations teams to align DevOps strategies with overall business objectives and ensure zero-downtime deployments.
- Ensure robust monitoring, alerting, change management, and incident response processes, leading root cause analysis (RCA) activities for system failures.
- Understand cloud networking, identity management, access controls, and compliance frameworks, to ensure integration and adherence to industry regulations and internal security policies.
- Oversee daily DevOps operations, establishing SLAs, SLIs, and SLOs aligned with operational goals to ensure system uptime and reliability.
- Establish and govern enterprise release practices, including environment management and SaaS deployments, driving modernization initiatives across cloud-native architectures.
- Identify opportunities for process improvements, automation, and innovation, monitoring trends, evaluating new tools and emerging technologies to enhance the DevOps ecosystem.
- Participate in long-term planning, manage capacity and resource allocation, and oversee vendor relationships and third-party service providers.
- Conduct regular performance reviews, provide feedback, and create development plans to support career growth and skill development.
- Other duties as assigned.
What you've accomplished:
- 8+ years of experience in software development, DevOps, or Site Reliability Engineering (SRE), and 5 years in a leadership or management role. At least 3 years in Dev/Ops with broad engineering experience preferred.
- Experience working with different technology stacks including successful modernization of legacy systems.
- Experience building and successfully leading Tier 1 applications, cloud transformation, resiliency strategies, and multi-site needs.
- Ability to define, measure and utilize metrics to communicate progress, track overall success and identify opportunities.
- Strong proficiency in software development methodologies, tools, and languages (e.g., Swift, SwiftUI, concurrency, JSON, REST, AI - Claude, Gemini, OpenAI, nice to have - Kotlin, Kotlin Multi Platform).
- Knowledge of DevOps practices, CI/CD pipelines, and automated testing.
- Excellent leadership, team-building, and interpersonal skills.
- Ability to motivate and inspire a diverse team of engineers.
- Strong problem-solving and decision-making abilities.
- Experience with agile development methodologies (e.g., Scrum, Kanban).
- Strong organizational skills and attention to detail.
- Ability to manage multiple projects and priorities simultaneously.
- Position is hybrid in-office Tuesday-Thursday but could require occasional in-office support on other days and up to 40% travel between offices.
- AWS Cloud Certification preferred.
- Flexible Time Off program
- Hybrid work model
- Variety of medical insurance options including a $0 cost premium employee coverage
- Benefits effective day 1 of employment include competitive 401K match with no vesting requirement, national health, dental, and vision plans
- Progressive family plan benefits
- An opportunity to work for a global and industry-leading technology company
- Impactful work in a collaborative environment