Sr Manager, Software Development & Engineering Senior

Charles Schwab Inc.Omaha, NE
$117,300 - $195,500Onsite

About The Position

As a Production Support Engineer, you will provide strategic and technical leadership for our Order Management System (OMS) in a high-availability environment. You will define reliability direction, lead enterprise-impact incident response, and establish scalable support practices that improve system resilience across teams.

Requirements

  • 4-year college degree in Computer Science or related field, or equivalent practical experience
  • 8+ years of experience in production support, SRE, software operations, or reliability engineering
  • Proven leadership in restoring and stabilizing mission-critical systems under high-pressure conditions
  • Expert-level troubleshooting in distributed systems, Java application stacks, and data-intensive platforms
  • Deep understanding of observability, incident command, and operational governance practices
  • Strong Oracle operational expertise and performance diagnostics background
  • Strong Linux platform knowledge (RHEL 7/8/9 preferred), including performance and capacity considerations
  • Advanced automation mindset with practical scripting and process design capabilities
  • Demonstrated success driving cross-team reliability initiatives and measurable operational outcomes
  • Executive-level communication skills for both technical and business stakeholders
  • Available on nights and weekends as required.

Responsibilities

  • Serving as the top technical escalation point for the most complex and business-critical production incidents
  • Setting reliability strategy, operational priorities, and service-level objectives for supported platforms
  • Leading cross-functional incident command and executive-facing communication during major disruptions
  • Defining and governing support standards, readiness criteria, and risk controls for production changes
  • Driving long-term reliability programs that reduce incident frequency and customer impact
  • Designing scalable operational models, including observability, alerting, and response frameworks
  • Partnering with architecture, engineering, database, and infrastructure leaders on platform direction
  • Evaluating systemic risks and making high-impact decisions under time pressure
  • Building team capability through mentorship, coaching, and technical leadership at scale
  • Establishing quality standards for incident analysis, problem management, and root-cause elimination
  • Influencing investment priorities based on reliability trends, capacity risks, and operational data
  • Ensuring production operations remain compliant, secure, and audit-ready
  • Supporting 24x7 operations on a rotating on-call schedule

Benefits

  • bonus or incentive opportunities
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service