Senior Site Reliability Engineer

Charles Schwab Inc.Southlake, TX
$130,000 - $155,000Onsite

About The Position

As a Site Reliability Engineer supporting the Cashiering organization, you will play a critical role in ensuring the stability, resiliency, and operational excellence of Schwab’s Move Money platforms. This role supports business-critical modernization initiatives, including Crypto and PRISM, while driving production readiness, availability, and continuous improvement across distributed applications. You will leverage observability, automation, and data-driven decision making to identify issues, improve system reliability, and optimize operational efficiency. Working closely with application development teams, infrastructure partners, business stakeholders, and external vendors, you will lead complex incident resolution efforts, influence resiliency strategies, and help shape operational best practices. Success in this role requires strong problem-solving skills, adaptability in a fast-paced environment, and the ability to build collaborative relationships while balancing reliability, performance, and business outcomes across mission-critical Cashiering services.

Requirements

  • 5+ years of experience supporting production applications, software engineering, site reliability engineering, DevOps, or related technology environments
  • Experience managing application reliability through Service Level Objectives (SLOs), monitoring practices, and operational health management
  • Demonstrated ability to troubleshoot and resolve complex issues across applications, platforms, infrastructure, and distributed systems
  • Experience with observability and monitoring tools including Splunk, Grafana, Datadog, AppDynamics, or ThousandEyes
  • Strong knowledge of Cashiering and Move Money functions including Wires, ACH, Journals, Checks, and Check Deposits
  • Experience reviewing and troubleshooting .NET and Java-based applications
  • Experience with AI-enabled productivity and automation tools such as GitHub Copilot and Microsoft Copilot Studio
  • Experience supporting CI/CD pipelines, production readiness processes, and operational support practices
  • Working knowledge of Jira, Bamboo, Confluence, Control-M, SQL Server, DB2, and IT service management disciplines
  • Strong analytical thinking, communication, collaboration, and problem-solving skills with the ability to coordinate effectively during critical incidents

Nice To Haves

  • Experience with Pivotal Cloud Foundry and Google Cloud Platform environments
  • Knowledge of Linux administration, including RHEL environments
  • Experience with shell scripting, Python, Perl, Ruby, or PowerShell
  • Understanding of disaster recovery, resiliency planning, and business continuity programs
  • Experience supporting distributed applications in highly regulated or financial services environments
  • Knowledge of network services including DNS, VPNs, load balancers, proxies, and firewalls
  • Experience working with SAN and NAS storage technologies
  • Experience mentoring team members and promoting operational standards and best practices

Responsibilities

  • Ensuring the stability, resiliency, and operational excellence of Schwab’s Move Money platforms.
  • Supporting business-critical modernization initiatives, including Crypto and PRISM.
  • Driving production readiness, availability, and continuous improvement across distributed applications.
  • Leveraging observability, automation, and data-driven decision making to identify issues, improve system reliability, and optimize operational efficiency.
  • Leading complex incident resolution efforts.
  • Influencing resiliency strategies.
  • Shaping operational best practices.

Benefits

  • bonuses or incentive opportunity
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service