Senior Manager, Support Engineering

OracleTX, TX
$98,700 - $209,500

About The Position

The Senior Manager, Repair Center Operations is responsible for building and scaling Oracle's centralized repair-center model for next-generation AI infrastructure. This role leads day-to-day repair operations and long-term capacity planning for Compute Tray (CT) repair, CDU validation, leak response, and regional expansion. Working across Development, Manufacturing, Supply Chain, Design, Architecture, and Global Product Engineering, the senior manager ensures repair centers are staffed, equipped, and ready to support routine demand as well as surge events while driving improvements in diagnosability, repair velocity, and serviceability. The role also serves as a key operational leader for the repair organization, translating product and infrastructure requirements into an executable staffing and facility plan. This includes establishing operating rhythms, tracking performance metrics, managing risk, and ensuring the repair network scales with current and future Oracle data center sites.

Requirements

  • Senior Manager, Repair Center Operations experience
  • Experience in building and scaling centralized repair-center models
  • Experience with next-generation AI infrastructure
  • Experience in leading day-to-day repair operations
  • Experience in long-term capacity planning for Compute Tray (CT) repair, CDU validation, leak response, and regional expansion
  • Experience working across Development, Manufacturing, Supply Chain, Design, Architecture, and Global Product Engineering
  • Experience ensuring repair centers are staffed, equipped, and ready to support routine demand and surge events
  • Experience driving improvements in diagnosability, repair velocity, and serviceability
  • Experience as a key operational leader for a repair organization
  • Experience translating product and infrastructure requirements into executable staffing and facility plans
  • Experience establishing operating rhythms, tracking performance metrics, and managing risk
  • Experience ensuring repair networks scale with current and future data center sites
  • Experience with GB200, GB300, and future AI compute platforms
  • Experience in capacity planning, shift coverage, resource allocation, and operational ramp for new sites
  • Experience defining repair-center requirements and resolving cross-functional issues
  • Experience overseeing repair workflows for Compute Trays, CDU validation, and leak-response activities
  • Experience establishing and managing operating metrics, reporting, and escalation paths
  • Experience driving standardization of repair-center processes, documentation, and best practices
  • Experience supporting regional scale-up by aligning repair-center buildouts with demand forecasts, site deployment schedules, and new platform introductions
  • Experience leading hiring, onboarding, development, and performance management
  • Experience managing competing demands and ensuring objectives are met on time and at expected quality level
  • Experience building strong cross-functional partnerships and maintaining alignment
  • Experience using data, trend analysis, and operational feedback to identify risks, improve performance, and recommend process enhancements
  • Experience maintaining a culture of accountability, safety, continuous improvement, and technical excellence
  • Experience developing scalable operating models for steady-state and surge repair work
  • Experience communicating clearly with senior leadership on site status, risks, dependencies, and required decisions

Responsibilities

  • Lead the planning, staffing, and execution of repair-center operations supporting GB200, GB300, and future AI compute platforms.
  • Own repair-center readiness, including capacity planning, shift coverage, resource allocation, and operational ramp for new sites.
  • Partner with Design, Architecture, Development, Manufacturing, Supply Chain, and Global Product Engineering to define repair-center requirements and resolve cross-functional issues.
  • Oversee repair workflows for Compute Trays, CDU validation, and leak-response activities to ensure safe, timely, and high-quality execution.
  • Establish and manage operating metrics, reporting, and escalation paths to maintain throughput, quality, and readiness across the repair network.
  • Drive standardization of repair-center processes, documentation, and best practices across current and future Oracle Repair locations.
  • Support regional scale-up by aligning permanent repair-center buildouts with demand forecasts, site deployment schedules, and new platform introductions.
  • Lead hiring, onboarding, development, and performance management for repair-center team members and supporting technical resources.
  • Set clear priorities, manage competing demands, and ensure repair-center objectives are met on time and at the expected quality level.
  • Build strong cross-functional partnerships and maintain alignment across engineering, operations, supply chain, and site leadership teams.
  • Use data, trend analysis, and operational feedback to identify risks, improve repair performance, and recommend process enhancements.
  • Maintain a culture of accountability, safety, continuous improvement, and technical excellence within the repair organization.
  • Translate business needs into executable plans for staffing, space, power, tooling, and repair capacity.
  • Develop scalable operating models that support both steady-state repair work and surge events without disrupting broader deployment goals.
  • Communicate clearly with senior leadership on site status, risks, dependencies, and required decisions.

Benefits

  • Medical, dental, and vision insurance, including expert medical opinion
  • Short term disability and long term disability
  • Life insurance and AD&D
  • Supplemental life insurance (Employee/Spouse/Child)
  • Health care and dependent care Flexible Spending Accounts
  • Pre-tax commuter and parking benefits
  • 401(k) Savings and Investment Plan with company match
  • Flexible Vacation
  • 11 paid holidays
  • 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
  • Paid parental leave
  • Adoption assistance
  • Employee Stock Purchase Plan
  • Financial planning and group legal
  • Voluntary benefits including auto, homeowner and pet insurance
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service