About The Position

As a member of CIBC’s Canadian Personal, Digital, Investing Technology and IT Transformation team, you’ll play a key role in enhancing the reliability, performance, and availability of complex software systems and services. As a Consultant, Site Reliability Engineering, you’ll apply advanced software engineering principles to design and implement robust monitoring frameworks, automate operational workflows, and optimize end-to-end client experiences. You’ll guide teams in balancing feature innovation with operational excellence, establish error budget-based decision-making frameworks, and champion a culture of continuous improvement. You’ll also lead incident response and root cause analysis, collaborate with cross-functional partners, and ensure compliance with evolving policies and standards to support enterprise-wide reliability goals. At CIBC we enable the work environment most optimal for you to thrive in your role. You'll have the flexibility to manage your work activities within a hybrid work arrangement where you'll spend 1-3 days per week on-site, while other days will be remote.

Requirements

  • Minimum of 5 years of experience in site reliability engineering, application support, or DevOps roles, with a strong focus on automation and operational excellence.
  • Demonstrated expertise in automating deployments, incident response workflows, and operational tasks using scripting and orchestration tools such as Python, PowerShell, Bash, or JavaScript.
  • Hands-on experience applying automation and machine learning to observability, anomaly detection, and incident resolution, and are proficient with tools like Dynatrace and Splunk.
  • Excel at collaborating across teams, building relationships, and sharing knowledge to drive reliability improvements, even when you do not have direct authority.
  • Familiarity with application security best practices, including authentication, authorization, secrets management, and secure coding principles.
  • A bachelor’s degree in computer science, engineering, or a related technical field, or possess equivalent practical experience.
  • Bring your real self to work, and you live our values - trust, teamwork, and accountability.

Responsibilities

  • Manage and optimize automated application deployments across multiple environments, ensuring consistency, reliability, and efficient releases.
  • Proactively monitor system health and performance, lead incident response and post-mortem analysis, and implement preventive measures to strengthen operational resilience.
  • Evaluate and enhance systems, processes, and tools to drive reliability, scalability, and efficiency, focusing on reducing manual operational work through automation.
  • Define, implement, and maintain observability strategies, including telemetry, logging, tracing, and alerting, to ensure deep visibility into application behavior and performance.
  • Consult and collaborate across the organization to align on reliability best practices, support consistency, and contribute to the development of standards and documentation.
  • Coordinate with production support teams to ensure operational readiness, analyze service level indicators and objectives, and maintain risk registries.

Benefits

  • competitive salary
  • incentive pay
  • banking benefits
  • a benefits program
  • defined benefit pension plan
  • an employee share purchase plan
  • a vacation offering
  • wellbeing support
  • MomentMakers, our social, points-based recognition program.
  • Purpose Day; a paid day off dedicated for you to use to invest in your growth and development.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service