About The Position

Apple's Information Systems and Technology (IS&T) organization is seeking a CDN Capacity & Infrastructure Manager to join the Infrastructure Services team. This role is crucial for managing data center equipment and systems that deliver compute, storage, and networking services for various teams at Apple. The Edge Services team, specifically, is responsible for the resiliency and quality of customer experiences through their hardware and software in global data centers. This manager will own capacity planning, forecasting, and infrastructure health for our global content delivery network (CDN), acting as a liaison between traffic trends, hardware procurement, and operational reliability. The goal is to ensure the edge infrastructure scales ahead of demand efficiently, balancing cost and utilization. This position is vital for supporting expanding edge delivery and high-concurrency traffic events like major OS releases, live sports/entertainment streaming, and continuous service growth. The role involves leading demand forecasting, fleet lifecycle execution, and unit-economics optimization for the CDN. It bridges edge software engineering, global network peering/transit, and the hardware supply chain, requiring the establishment of rigorous capacity models, mitigation of long-lead hardware risks, and optimization of multi-terabit peering commitments. By transitioning the edge footprint to predictive, data-driven modeling, this role directly protects user experience and drives discipline into infrastructure CapEx spend.

Requirements

  • 6+ years of experience in capacity planning, infrastructure engineering, or network operations, ideally within CDN, cloud, or large-scale distributed systems environments
  • Experience forecasting infrastructure capacity using traffic/usage data
  • Understanding of CDN architecture fundamentals (edge/PoP design, caching, traffic routing, origin shielding)
  • Experience working with monitoring and analytics tooling to track utilization and performance metrics
  • Track record coordinating cross-functionally with network engineering, SRE, and vendor/procurement teams
  • Data analysis skills (SQL and/or scripting) to build forecasting models and reporting
  • Written and verbal communication skills, including presenting technical findings to leadership
  • Bachelor’s degree in CS, EE, or related field, or equivalent practical experience

Nice To Haves

  • Familiarity with time-series forecasting methods or capacity modeling tools
  • Experience with data center or colocation site planning and hardware lifecycle management
  • Background in live-streaming or high-traffic event capacity planning
  • Experience with observability stacks (e.g., Splunk, ClickHouse, Grafana) for capacity/utilization reporting
  • Exposure to automation/orchestration tooling for infrastructure provisioning

Responsibilities

  • Own end-to-end capacity planning for CDN edge and origin infrastructure, including short-term (weeks), medium-term (quarters), and long-term (annual) forecasting models
  • Analyze traffic patterns, bandwidth utilization, and growth trends across PoPs/edge nodes to identify capacity risks
  • Partner with network engineering and SRE teams to plan and execute hardware deployments, decommissions, and capacity rebalancing across regions
  • Build and maintain dashboards/reporting that track utilization, headroom, and forecast accuracy against actuals
  • Lead capacity readiness reviews ahead of major traffic events (product launches, live streams, seasonal peaks)
  • Coordinate with vendors and colocation partners on lead times, hardware procurement, and site buildouts
  • Develop and refine capacity models that account for traffic imbalance, failover scenarios, and regional redundancy requirements
  • Drive root-cause analysis on capacity-related incidents and feed learnings back into forecasting models
  • Collaborate with finance/procurement on budget planning tied to infrastructure growth
  • Establish and maintain runbooks, SOPs, and escalation paths for capacity-related operational issues
  • Present capacity status, risks, and recommendations to engineering leadership on a recurring cadence
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service