Senior Systems Operations Engineer

Wells Fargo & CompanyIrving, TX
Onsite

About The Position

Wells Fargo is seeking a Senior Systems Operation Engineer within the Consumer Technology (CT) organization, providing technology solutions to the LOBs and managing the application portfolios through the enablers of skills, stability, security, scalability, speed, and success. We are seeking a highly motivated Senior Engineer in Platform Services to drive operational excellence, reliability, and support readiness across a portfolio of data products and platforms. This role combines Site Reliability Engineering (SRE) practices and cross-functional collaboration to ensure stable, resilient, and supportable solutions.

Requirements

  • 4+ years of Systems Engineering, Technology Architecture experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education
  • 3+ years of Strong experience in production support, SRE, or platform engineering in enterprise environments
  • 2+ years of Hands-on experience with observability tools (Grafana, Splunk, GCP monitoring tools)
  • 2+ years of experience with distributed systems, data pipelines, event-driven architectures preferably on Google Cloud Platform (Airflow)

Nice To Haves

  • Proven ability to lead high-severity incidents with effective stakeholder communication and executive reporting.
  • Strong collaboration with cross-functional teams (AppDev, Infra, Business stakeholders)
  • Ability to operate in high-pressure incidents with clear communication
  • Google Data Engineer certification.
  • Splunk certification.
  • Familiarity with CI/CD and Infrastructure as Code.
  • Demonstrate strong ownership, accountability, and a customer-first mindset.
  • Proactively identify opportunities to improve reliability, supportability, and operational efficiency.

Responsibilities

  • Drive observability maturity across applications (metrics, logging, alerting, dashboards)
  • Lead incident response activities, including triage, stakeholder communication, resolution coordination, and Root Cause Analysis (RCA) completion.
  • Drive problem management including trend analysis and permanent fix tracking
  • Support application support through participation in on-call rotation.
  • Champion operational excellence through automation, process improvements, and reduction of manual support activities.
  • Partner with AppDev teams to ensure runbook quality, support readiness, and operational documentation completeness.
  • Mentor team members and drive cross-training initiatives to improve team coverage, resiliency, and technical depth.

Benefits

  • Relocation assistance is not available for this position.
  • Visa Sponsorship is not available for this position.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service