- Salary
- $163k – $198k
- Location
- USA - FL - Kirkman Point 2, United States of America
- Workplace
- Onsite
- Type
- Full-time
- Department
- Engineering
- Seniority
- Manager
- Experience
- 8+ years
- Closing date
- Today
- Source
- Workday
Description
Job Posting Title:
Manager Site Reliability EngineeringReq ID:
10156470Job Description:
This is not a remote role. You must be in the local area or be willing to relocate
“We Power the Magic!” That’s our motto at Disney Experiences (DX). Our team creates world-class immersive digital experiences for the Company’s premier vacation brands including Disney’s Parks & Resorts worldwide, Disney Cruise Line, Aulani, a Disney Resort & Spa, and Disney Vacation Club.
We are responsible for the end-to-end digital and physical Guest experience for all technology & digital-led initiatives across the Attractions & Entertainment, Food & Beverage, Resorts & Transportation and Merchandise lines of business as well as other initiatives including MyDisneyExperience and Hey, Disney!
This role sits in the Commerce Site Reliability Engineering organization within Technology & Digital for Disney Experiences. It works closely with Commerce and Consumer Products personnell from across the company.
About The Role & Team:
- The Commerce Site Reliability Engineering Manager supports the Parks Commercial Systems and Consumer Products systems
- The manager is responsible for a team that performs incident management and requests for work in both the Systems Engineering and Site Reliability engineering functions
- The manager will be working closely to manage a diverse team of engineers to deliver observability, availability and security to the application teams in both spaces
- The team also includes working closely with other leaders to deliver functions for the overall commerce area including SDLC functions and AI functions
- The leader will be responsible for managing a team of Cast Members including regular reviews and OKR’s.
What You'll Do:
- Oversee finances and budgets in MyPPM, ensure accurate billing processes, and contribute to forecasting and accrual processes to maintain financial integrity and support organizational objectives
- Lead the evolution of DevOps practices within the broader team framework, guiding others in leveraging this culture to enhance observability practices
- Manage the Site Reliability Engineers to deliver monitoring and observability for the development and business users as needed
- Manage the design, build, and support of products platforms
- Drive teams to consult, design, build, and support development pipelines, automate infrastructure and operations, build telemetry for monitoring, engineer high-reliability and reinforce best-practices to secure company data
- Lead all aspects of systems administration skills on Google, Amazon and Azure clouds as well as on premise systems and must have extensive experience with web technologies, source control management using Harness and Git.
- Strategize systems administration in RHEL, Bottle rocket, Kubernetes and containers and bring knowledge on systems, network, operational excellence and application stability, security, performance, and capacity management, operational excellence and application stability, security, performance, and capacity management, as well as documentation
- Engage in estimation and planning across the organization, voicing recommendations, feedback, and solutions from a technical perspective and aligning to the overall project goals to deliver on-time & in-scope
- Proactively track and assess new technologies across the industry to inform strategic decision-making and recommendations
Required Qualifications:
- Must have a minimum 8 years of related work experience
- Demonstrated leadership in implementing observability principles across complex systems and environments, fostering a culture of reliability and resilience
- Extensive experience with a wide range of continuous integration tools, including Gitlab, AWS CodeBuild, CodeDeploy, CodePipeline, and Azure DevOps, optimizing workflows and ensuring seamless deployment processes
- Proficiency in designing and managing highly scalable and resilient infrastructure using configuration management and orchestration tools such as Terraform, Cloud Formation, Ansible, and Chef, driving operational excellence and efficiency
- Leveraging AI for predictive insights, driving continuous improvement in system reliability
- Outstanding communication and leadership abilities, to ensure effective growth and development of team
- A visionary who motivates teams to excel and fosters creativity, consistently driving excellence in all endeavors
- An advocate for a diverse and inclusive culture that encourages innovation and ensures every team member feels a sense of belonging
Preferred Qualifications:
- A deep understanding of containerized and serverless architectures and strategies
- Experience with Kubernetes
- Experience with AWS and / or GCP
- Experience with management of a cross functional team
Required Education:
- Bachelor’s degree in Computer Science, Information Systems, Software, Electrical or Electronics Engineering, or comparable field of study, and/or equivalent work experience
#DISNEYTECH
The hiring range for this position in Orlando is $163,400.00 - $198,000.00 per year. The base pay actually offered will take into account internal equity and also may vary depending on the candidate’s geographic region, job-related knowledge, skills, and experience among other factors. A bonus and/or long-term incentive units may be provided as part of the compensation package, in addition to the full range of medical, financial, and/or other benefits, dependent on the level and position offered.Job Posting Segment:
DX TechnologyJob Posting Primary Business:
Tech Delivery, Platforms, & Core SystemsPrimary Job Posting Category:
Site/System Reliability EngineerEmployment Type:
Full timePrimary City, State, Region, Postal Code:
Orlando, FL, USAAlternate City, State, Region, Postal Code:
Date Posted:
2026-08-04