Lead Site Reliability Engineer
Top focus
Job Posting Title: Lead Site Reliability Engineer Req ID: 10155442 Job Description: About The Role & Team “We Power the Magic!” That’s our motto at Disney Experiences (DX). Our team creates world-class immersive digital experiences for the Company’s premier vacation brands including Disney’s Parks & Resorts worldwide, Disney Cruise Line, Aulani, a Disney Resort & Spa, and Disney Vacation Club.
We are responsible for the end-to-end digital and physical Guest experience for all technology & digital-led initiatives across the Attractions & Entertainment, Food & Beverage, Resorts & Transportation and Merchandise lines of business as well as other initiatives including MyDisneyExperience and Hey, Disney!
The US Parks Site Reliability Organization is accountable for the reliability and resilience of a portfolio of critical applications and services. We partner with product, engineering, and SRE teams across the organization to define and evolve reliability standards rooted in SRE and DevOps principles.
Rather than simply operating systems, we apply engineering approaches—observability, automation, and data ‑ driven decision making—to proactively improve service health. By designing reliability into platforms and reducing operational toil, we enable teams to focus on innovation and delivering world ‑ class guest experiences.
This role sits in the US Parks & Resorts organization within Disney Experiences Technology and works closely with other site reliability engineers, application delivery teams, and systems engineers from across the company. The Lead Site Reliability Engineer will report to the Manager, System Engineering .
What You Will Do Serve as the SRE subject matter expert and technical lead for assigned products and platforms, owning the reliability strategy and embedding SRE and DevOps best practices Drive adoption and tracking of service level management—defining and operationalizing SLIs, SLOs, and SLAs—for the systems and applications in your assigned portfolio Lead the design, build, and support of products and platforms; consult on and build development pipelines, automate infrastructure and operations, and create telemetry for monitoring Engineer high reliability and reinforce best practices to secure company data across systems, network, performance, capacity, and operational excellence Mentor and guide other site reliability and systems engineers, providing coaching, feedback, and technical direction to elevate team performance and hold self and others accountable to commitments Lead Major Incident response for owned services—minimizing Mean Time to Resolve and delivering comprehensive retrospectives that result in measurable improvements to prevent future failures Partner with engineering, product, and program management to align priorities, manage dependencies, contribute to estimation and planning, and negotiate solutions to complex reliability challenges Champion a DevOps culture and a shift-left, reliability-by-design mindset among peers and developers Stay current with emerging technologies and apply AI/automation to reduce toil and improve service health Required Qualifications & Skills Minimum 7 years of related work experience Proficient in agile environments Applied expertise in observability principles and tools, including defining and implementing SLIs, SLOs, and SLAs Hands-on experience with CI/CD tools like Gitlab, AWS CodeBuild , Azure DevOps Proficient in configuration management tools: Terraform, CloudFormation, Ansible, Chef Experience in procedural programming languages ( Python, Perl, Ruby, Java, Go, Rust, C/C++ ) Skilled in Cloud environments ( AWS, Azure, Google Cloud ) Proven ability to design and build reliable, scalable enterprise systems Capable of leading reliability efforts and identifying root causes in large-scale distributed systems Proficient in UNIX/Linux administration, troubleshooting, and security Demonstrated experience leading technical projects and ensuring smooth delivery Collaborative work with Security Operations teams for secure solutions Strong troubleshooting skills across systems, network, and code Proven experience mentoring, guiding, or training other engineers, with strong written and verbal communication and the ability to influence without direct authority Proactive demeanor toward continuous learning and mastering emerging tools and methodologies Education Bachelor’s degree in Computer Science , Information Systems, Software, Electrical or Electronics Engineering, or comparable field of study, and/or equivalent work experience required The hiring range for this position in Florida is $148,300 to $198,800 per year.
The base pay actually offered will take into account internal equity and also may vary depending on the candidate’s geographic region, job-related knowledge, skills, and experience among other factors. A bonus and/or long-term incentive units may be provided as part of the compensation package, in addition to the full range of medical, financial, and/or other benefits, dependent on the level and position offered.
Job Posting Segment: DX Technology Job Posting Primary Business: US Parks Primary Job Posting Category: Site/System Reliability Engineer Employment Type: Full time Primary City, State, Region, Postal Code: Bay Lake, FL, USA Alternate City, State, Region, Postal Code: Date Posted: 2026-07-31