RHEL Systems Administrator/DevOps Engineer

BMOMississauga, ON
CA$75,900 - CA$141,900Hybrid

About The Position

We are looking for a highly skilled DevOps / Site Reliability Engineer (SRE)/Sys Admin to join one of our high-performing, mission-critical technology teams. This role offers the opportunity to work on premium financial platforms, driving reliability, automation, and innovation at scale. As part of our DevOps/SRE team, you will play a key role in designing, building, and operating resilient, secure, and scalable infrastructure that supports our critical business applications.

Requirements

  • 5+ years of experience in DevOps, SRE and Systems Administration or related roles in hybrid (on-prem and AWS) environments
  • Strong experience with RHEL systems administration and clustering technologies (e.g., Veritas Cluster)
  • Hands-on experience with AWS IaaS and Infrastructure as Code (CDK / TypeScript preferred)
  • Hands-on experience with configuration management tools (Ansible, YAML)
  • Strong experience with observability platforms such as Dynatrace and CloudWatch
  • Proficiency in scripting and development using Python, Bash, and/or JavaScript
  • Experience implementing automation-first solutions across infrastructure and application layers
  • Solid understanding of security and compliance practices within regulated industries (financial services preferred)
  • Experience with Git-based workflows (GitHub preferred)
  • Working knowledge of ServiceNow and ITSM processes (Incident, Problem, Change, Release, Configuration Management)
  • Proven ability to support and operate large-scale, mission-critical systems
  • Strong analytical and problem-solving abilities
  • Effective communication and documentation skills
  • Ability to collaborate across multiple teams and stakeholders
  • Demonstrated ability to coach and mentor team members
  • Data-driven decision-making mindset
  • Deep understanding of IT operational processes, monitoring, logging, and alerting standards

Nice To Haves

  • CDK / TypeScript preferred
  • financial services preferred
  • GitHub preferred

Responsibilities

  • Partner with development, operations, and security teams to design and deliver secure, scalable, and resilient infrastructure solutions
  • Build and maintain automation frameworks for deployment, scaling, and observability
  • Design and implement CI/CD pipelines, release strategies, and recovery mechanisms
  • Continuously improve system performance, availability, reliability, and security posture
  • Provide end-to-end ownership of mission-critical platforms, including production support and root-cause analysis
  • Proactively monitor systems using observability tools to identify and address performance and reliability improvements
  • Lead or contribute to incident response, triage, and resolution with a focus on rapid recovery and prevention
  • Support deployment activities and manage implementation issues through to resolution
  • Drive adoption of modern engineering practices, tools, and processes to enhance delivery and operational efficiency
  • Analyze complex technical issues and recommend solutions aligned with business impact
  • Ensure compliance with enterprise standards and regulatory requirements
  • Participate in an on call rotation to support production systems (if needed)

Benefits

  • health insurance
  • tuition reimbursement
  • accident and life insurance
  • retirement savings plans
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service