Site Reliability Engineer

PNC BankPhoenix, AZ
Onsite

About The Position

At PNC, our people are our greatest differentiator and competitive advantage in the markets we serve. We are all united in delivering the best experience for our customers. We work together each day to foster an inclusive workplace culture where all of our employees feel respected, valued and have an opportunity to contribute to the company’s success. As a Site Reliability Engineer within PNC's Technology organization, you can be based in Pittsburgh PA, Strongsville OH, Birmingham AL, Denver CO, Phoenix CO or Dallas TX. PNC will not provide sponsorship for employment visas or participate in STEM OPT for this position. We are seeking a Site Reliability Engineer to join our Site Reliability Center, supporting the continuous improvement of critical retail banking applications. This role focuses on developing monitoring solutions, operational dashboards, and data-driven insights that enhance application reliability, performance, and observability across the enterprise. The ideal candidate has a blend of data analytics, dashboard development, and software operations experience, with a passion for transforming operational data into actionable insights. This individual will partner with engineering and support teams to build and maintain tools that improve service reliability and operational efficiency.

Requirements

  • Bachelor's degree in Data Science, Computer Science, Information Technology, or a related field preferred or equivalent experience
  • Experience building dashboards and data visualizations using Grafana or similar business intelligence/monitoring platforms.
  • Understanding of application monitoring, observability, and software operations concepts.
  • Experience analyzing and interpreting large data sets to support operational decision-making.
  • Strong proficiency with Microsoft Excel, reporting, and data presentation.
  • Ability to manage multiple projects and work effectively in a highly collaborative environment.
  • Strong communication and stakeholder management skills.
  • Roles at this level typically require a university / college degree, with 2+ years of relevant professional experience. In lieu of a degree, a comparable combination of education, job specific certification(s), and experience (including military service) may be considered.

Nice To Haves

  • Experience with performance monitoring and observability platforms such as: Dynatrace, Elastic, Splunk, Similar enterprise monitoring solutions
  • Working knowledge of Python and/or Java.
  • Experience supporting Site Reliability Engineering (SRE), Operations, or Application Support organizations.
  • Familiarity with retail banking or financial services environments.
  • Experience developing automated reporting and monitoring solutions.

Responsibilities

  • Design, build, and maintain operational and service reliability dashboards using Grafana and related visualization tools.
  • Support ongoing continuous improvement initiatives for retail banking applications.
  • Collaborate with cross-functional technology teams to gather requirements and develop reporting and monitoring solutions.
  • Analyze application and operational data to identify trends, performance issues, and opportunities for optimization.
  • Develop and maintain observability solutions that provide visibility into application health, performance, and reliability.
  • Track and manage dashboard-related projects and initiatives from request intake through completion.
  • Create reports, presentations, and metrics that communicate operational performance to technical and business stakeholders.
  • Partner with Site Reliability and Engineering teams to improve monitoring strategies and operational processes.
  • Contribute to the development and maintenance of internal tools supporting application monitoring and performance management.
  • Designs and develops systems that are resilient and highly performant at tremendous scale.
  • Partners to develop engineering toolset to drive stability & automation.
  • Assesses opportunities to drive engineering stability through the analytics and metrics.
  • Develops and implements automated and sustainable monitoring and alerting sites to ensure the availability and performance of critical applications.
  • Collaborates cross-functionally to gather and analyze metrics from operating sites and applications to assist in performance tuning and fault finding as well as improve scalability and reliability metrics.
  • Troubleshoots incidents and participates in testing approaches and test strategy results.
  • Performs analytics on previous incidents and usage patterns to better predict issues and take proactive actions.
  • Identifies opportunities to evangelize adoption for greater self-healing and resiliency patterns.
  • Provides operational support and engineering for multiple large-scale distributed software applications.
  • Partners with technology teams across the enterprise to establish Site Reliability Engineering best practices and automated solutions with a focus on operational excellence.

Benefits

  • medical/prescription drug coverage (with a Health Savings Account feature)
  • dental and vision options
  • employee and spouse/child life insurance
  • short and long-term disability protection
  • 401(k) with PNC match
  • pension and stock purchase plans
  • dependent care reimbursement account
  • back-up child/elder care
  • adoption, surrogacy, and doula reimbursement
  • educational assistance, including select programs fully paid
  • a robust wellness program with financial incentives
  • maternity and/or parental leave
  • up to 11 paid holidays each year
  • 9 occasional absence days each year, unless otherwise required by law
  • between 15 to 25 vacation days each year, depending on career level; and years of service
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service