Sr Principal Software Engineer

The Walt Disney CompanySeattle, WA
$239,700 - $321,400Onsite

About The Position

Disney Entertainment and ESPN Product & Technology is a global organization of engineers, product developers, designers, technologists, data scientists, and more – all working to build and advance the technological backbone for Disney’s media business globally. The team marries technology with creativity to build world-class products, enhance storytelling, and drive velocity, innovation, and scalability for our businesses. We are Storytellers and Innovators. Creators and Builders. Entertainers and Engineers. We work with every part of The Walt Disney Company’s media portfolio to advance the technological foundation and consumer media touch points serving millions of people around the world. Building the future of Disney’s media: Our Technologists are designing and building the products and platforms that will power our media, advertising, and distribution businesses for years to come. Reach, Scale & Impact: More than ever, Disney’s technology and products serve as a signature doorway for fans’ connections with the company’s brands and stories. Disney+. Hulu. ESPN. ABC. ABC News…and many more. These products and brands – and the unmatched stories, storytellers, and events they carry – matter to millions of people globally. Innovation: We develop and implement groundbreaking products and techniques that shape industry norms, and solve complex and distinctive technical problems. Product Engineering is a unified team responsible for the engineering of Disney Entertainment & ESPN digital and streaming products and platforms. This includes product engineering, media engineering, quality assurance, engineering behind personalization, commerce, lifecycle, and identity. Within Product Engineering, the Reliability Engineering organization’s mission is to make the complex invisible: every developer at Disney+, ESPN, and Hulu should be able to build, ship, and operate great software without having to solve foundational distributed systems problems from scratch. We operate across four capability areas: platform services, site reliability engineering, observability, and AI operations. This role sits at the center of the platform services capability. It is one of the most consequential technical bets we are making to improve engineering velocity and reliability across 4,000 engineers.

Requirements

  • 12+ years of software engineering experience, with at least 5 years in a Staff, Principal, or higher‑level individual contributor role owning large‑scale systems.
  • Direct experience designing, building, and operating platform or infrastructure services in an organization of 1,000+ engineers, including navigating cross‑org coordination at hyperscale.
  • Deep expertise in distributed systems and platforms, such as API gateways, service mesh, distributed tracing, observability platforms, rate limiting, circuit breaking, and load shedding.
  • Proven track record of delivering platform services as software that achieve broad internal adoption; you understand what to build so engineers choose to use the platform.
  • Hands‑on experience with service‑to‑service authentication, zero‑trust networking, and identity federation at scale.
  • Demonstrated ability to communicate technical architecture and platform strategy to Directors and VPs, translating system design decisions into measurable velocity, cost, and reliability outcomes.
  • Experience leading teams through build‑versus‑buy decisions for platform‑level capabilities, including prototyping, vendor evaluation, and long‑term ownership considerations.
  • Related Bachelor’s Degree or Higher

Nice To Haves

  • Experience at a hyperscaler or large consumer technology company (e.g., Google, Meta, Netflix, Apple, Microsoft, or equivalent) with platform engineering, and small company experience as well.
  • Prior work building infrastructure that AI or ML systems depend on, such as inference routing, feature stores, model telemetry, or experiment tracking.
  • Background in streaming media infrastructure: CDN integration, live event architecture, or high‑concurrency content delivery systems.
  • Experience defining and governing SLO frameworks, error budgets, and production readiness standards across a large engineering organization.

Responsibilities

  • Design, implement, and own the technical architecture and core software components for Disney’s platform services layer, including: API gateway and edge services, Service‑to‑service authentication and authorization, Rate limiting, load shedding, and circuit breaking, Distributed tracing, telemetry capture, and logging pipelines, User identity sharing and cross‑service context propagation, Observability and experimentation platforms.
  • Write, review, and ship production‑grade code for shared libraries, services, SDKs, CLI tools, and golden‑path templates that are adopted by 400+ teams.
  • Produce reference implementations and starter kits that embody platform standards and patterns, so teams can adopt best practices by default.
  • Produce architecture decision records (ADRs), technical design docs, and platform RFCs that guide delivery teams and establish durable institutional standards.
  • Define and maintain golden path templates and tooling that make the “correct” implementation the easiest path for teams consuming platform services.
  • Establish API and event schema standards, SLO frameworks, error budgets, and upgrade‑path contracts across the platform services portfolio, ensuring compatibility, backward‑compatibility where appropriate, and smooth migrations.
  • Review and approve service designs and platform integrations before they enter the build phase, resolving cross‑domain technical dependencies before they become delivery blockers.
  • Partner with Directors of Platform Engineering and Site Reliability to translate architecture into executable quarterly roadmaps, with clear milestones and ownership for software delivery.
  • Evaluate build‑versus‑buy decisions with rigor, including prototyping, integration spikes, and cost modeling; quantify total cost of ownership, integration risk, and long‑term maintainability.
  • Lead deep technical investigations into complex production issues across distributed systems, including cross‑service performance bottlenecks, cascading failures, and resiliency gaps.
  • Design and implement platform‑level resiliency patterns (e.g., retries, bulkheads, fallback strategies) into shared libraries and services so teams inherit reliability by default.
  • Partner with SRE and observability teams to ensure platform services have excellent telemetry, dashboards, and alerting, and that incidents drive improvements to the platform itself.
  • Ensure platform primitives are AI‑ready from day one: design infrastructure that AI and ML systems can rely on, including support for: Inference routing and traffic splitting, Model telemetry and monitoring, Experiment tracking and feature‑flagging for models and AI‑powered experiences.
  • Collaborate with AI/ML teams to build reusable software components and patterns that accelerate AI workloads on the platform.
  • Operate as an executive‑level individual contributor (P6): drive technical direction across multiple organizations, influence Directors and VPs, and make decisions that materially shape the platform strategy.
  • Mentor Principal and Senior Engineers across the organization through pairing, code reviews, design critiques, and architecture reviews, building the next generation of engineering leaders.
  • Communicate platform service strategy, roadmap progress, and key technical trade‑offs to domain engineering Directors and VPs, translating architecture decisions into velocity, cost, and reliability outcomes for the business.

Benefits

  • A bonus and/or long-term incentive units may be provided as part of the compensation package, in addition to the full range of medical, financial, and/or other benefits, dependent on the level and position offered.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service