Skip to content
Skip to content
DevOps Jobs
The Walt Disney Company

Lead Site Reliability Engineer

The Walt Disney Company

Location
Hybrid (USA - FL - Disney's Hollywood Studios - Feature Animation Building, Florida)
Compensation
$148k - $198k/yr
Employment
Full-time
Level
Senior Level
Posted 2 days ago

About the Role

Disney Experiences Technology seeks a Lead Site Reliability Engineer to own reliability strategy for US Parks & Resorts digital products. The role drives SRE best practices, automates infrastructure, and mentors engineers to enhance service health and guest experience.

Skills

Site Reliability Engineering DevOps Observability SLI SLO SLA Terraform Python AWS Azure CI/CD Linux Automation Incident Response System Engineering Cloud Infrastructure

Benefits

  • Medical insurance
  • Financial benefits

Perks

  • Bonus
  • Long-term incentive units

Full job details

Job Posting Title:

Lead Site Reliability Engineer

Req ID:

10155442

Job Description:

About The Role & Team

“We Power the Magic!” That’s our motto at Disney Experiences (DX). Our team creates world-class immersive digital experiences for the Company’s premier vacation brands including Disney’s Parks & Resorts worldwide, Disney Cruise Line, Aulani, a Disney Resort & Spa, and Disney Vacation Club. We are responsible for the end-to-end digital and physical Guest experience for all technology & digital-led initiatives across the Attractions & Entertainment, Food & Beverage, Resorts & Transportation and Merchandise lines of business as well as other initiatives including MyDisneyExperience and Hey, Disney! 

 

The US Parks Site Reliability Organization is accountable for the reliability and resilience of a portfolio of critical applications and services. We partner with product, engineering, and SRE teams across the organization to define and evolve reliability standards rooted in SRE and DevOps principles. Rather than simply operating systems, we apply engineering approaches—observability, automation, and datadriven decision making—to proactively improve service health. By designing reliability into platforms and reducing operational toil, we enable teams to focus on innovation and delivering worldclass guest experiences. 

This role sits in the US Parks & Resorts organization within Disney Experiences Technology and works closely with other site reliability engineers, application delivery teams, and systems engineers from across the company.   

The Lead Site Reliability Engineer will report to the Manager, System Engineering. 

 

What You Will Do

  • Serve as the SRE subject matter expert and technical lead for assigned products and platforms, owning the reliability strategy and embedding SRE and DevOps best practices
  • Drive adoption and tracking of service level management—defining and operationalizing SLIs, SLOs, and SLAs—for the systems and applications in your assigned portfolio
  • Lead the design, build, and support of products and platforms; consult on and build development pipelines, automate infrastructure and operations, and create telemetry for monitoring
  • Engineer high reliability and reinforce best practices to secure company data across systems, network, performance, capacity, and operational excellence
  • Mentor and guide other site reliability and systems engineers, providing coaching, feedback, and technical direction to elevate team performance and hold self and others accountable to commitments
  • Lead Major Incident response for owned services—minimizing Mean Time to Resolve and delivering comprehensive retrospectives that result in measurable improvements to prevent future failures
  • Partner with engineering, product, and program management to align priorities, manage dependencies, contribute to estimation and planning, and negotiate solutions to complex reliability challenges
  • Champion a DevOps culture and a shift-left, reliability-by-design mindset among peers and developers
  • Stay current with emerging technologies and apply AI/automation to reduce toil and improve service health 

 

Required Qualifications & Skills

  • Minimum 7 years of related work experience
  • Proficient in agile environments 
  • Applied expertise in observability principles and tools, including defining and implementing SLIs, SLOs, and SLAs 
  • Hands-on experience with CI/CD tools like Gitlab, AWS CodeBuild, Azure DevOps 
  • Proficient in configuration management tools: Terraform, CloudFormation, Ansible, Chef 
  • Experience in procedural programming languages (Python, Perl, Ruby, Java, Go, Rust, C/C++) 
  • Skilled in Cloud environments (AWS, Azure, Google Cloud) 
  • Proven ability to design and build reliable, scalable enterprise systems 
  • Capable of leading reliability efforts and identifying root causes in large-scale distributed systems 
  • Proficient in UNIX/Linux administration, troubleshooting, and security 
  • Demonstrated experience leading technical projects and ensuring smooth delivery 
  • Collaborative work with Security Operations teams for secure solutions 
  • Strong troubleshooting skills across systems, network, and code 
  • Proven experience mentoring, guiding, or training other engineers, with strong written and verbal communication and the ability to influence without direct authority 
  • Proactive demeanor toward continuous learning and mastering emerging tools and methodologies 

 

Education

  • Bachelor’s degree in Computer Science, Information Systems, Software, Electrical or Electronics Engineering, or comparable field of study, and/or equivalent work experience required
The hiring range for this position in Florida is $148,300 to $198,800 per year. The base pay actually offered will take into account internal equity and also may vary depending on the candidate’s geographic region, job-related knowledge, skills, and experience among other factors. A bonus and/or long-term incentive units may be provided as part of the compensation package, in addition to the full range of medical, financial, and/or other benefits, dependent on the level and position offered.

Job Posting Segment:

DX Technology

Job Posting Primary Business:

US Parks

Primary Job Posting Category:

Site/System Reliability Engineer

Employment Type:

Full time

Primary City, State, Region, Postal Code:

Bay Lake, FL, USA

Alternate City, State, Region, Postal Code:

Date Posted:

2026-07-31