Site Reliability Engineer - Cloud & Platform Engineering, Manulife Bank Technology

ManulifeWaterloo, ON
CA$86,100 - CA$136,100Hybrid

About The Position

Join Manulife Bank Technology team as a Site Reliability Engineer (SRE) and help deliver reliable, secure, and scalable technology services that support critical banking operations. In this role, you'll collaborate with engineering and operations teams to improve platform resilience, enhance observability, automate operational processes, and drive continuous service improvement. As part of the Service Delivery Management team, you will contribute to building and supporting highly available systems while applying modern Site Reliability Engineering practices. You'll have opportunities to expand your expertise in cloud technologies, automation, incident management, and operational excellence, helping ensure reliability remains a core feature of the platforms and services our customers depend on every day.

Requirements

  • Minimum 2-5 years+ of experience in Site Reliability Engineering, DevOps, Platform Engineering, Cloud Operations, or a related technology field.
  • Experience supporting applications and services in cloud environments such as Azure, AWS, or Google Cloud Platform.
  • Experience developing automation and scripting solutions using languages such as Python, Bash, PowerShell, or similar technologies.
  • Experience with monitoring, logging, and observability platforms such as New Relic, Grafana, Azure Data Explorer (ADX), or comparable tools.
  • Strong troubleshooting, analytical, and problem-solving skills, with the ability to effectively manage competing priorities and respond during service disruptions.
  • Excellent communication and collaboration skills, with the ability to work effectively across technical and business teams.

Nice To Haves

  • Experience working with container technologies such as Docker and Kubernetes.
  • Familiarity with ITIL practices and Agile delivery methodologies.
  • Experience creating operational dashboards and reporting using Power BI or similar analytics tools.
  • Experience working within financial services, banking, or other highly regulated environments.

Responsibilities

  • Design, build, and maintain infrastructure, automation, and tooling that improve system reliability, availability, scalability, and operational efficiency.
  • Partner with engineering teams to embed reliability, observability, and performance best practices throughout the software development lifecycle.
  • Monitor production systems and services, proactively identifying, troubleshooting, and resolving issues to minimize customer impact.
  • Develop and enhance observability capabilities, including metrics, logging, tracing, dashboards, and alerting solutions.
  • Participate in incident response and on-call support for critical systems and services.
  • Conduct root cause analyses following incidents and drive corrective and preventative actions to improve long-term platform reliability.
  • Define, track, and continuously improve Service Level Objectives (SLOs), Service Level Indicators (SLIs), and operational performance metrics.
  • Support disaster recovery, resilience testing, and business continuity initiatives to ensure service readiness and recovery capabilities.
  • Promote proactive monitoring, automation, and operational excellence practices that improve service reliability and reduce manual intervention.
  • Collaborate with technical and business partners to continuously improve service performance, stability, and customer experience.

Benefits

  • health
  • dental
  • mental health
  • vision
  • short- and long-term disability
  • life and AD&D insurance coverage
  • adoption/surrogacy and wellness benefits
  • employee/family assistance plans
  • various retirement savings plans (including pension and a global share ownership plan with employer matching contributions)
  • financial education and counseling resources
  • generous paid time off program in Canada includes holidays, vacation, personal, and sick days
  • full range of statutory leaves of absence
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service