Expert Observability Engineer

FinastraMississauga, ON
CA$112,000 - CA$130,000Hybrid

About The Position

At Finastra, we’re a global leader in financial services software, dedicated to expanding access to financial services and shaping what’s next for the industry. Our technology powers mission‑critical solutions across Lending, Payments and Universal Banking, supporting over 7,000 customers, including 80% of the world’s top 50 banks, in more than 110 countries. We are seeking a highly skilled and experienced Observability Engineer to lead and evolve our monitoring and observability capabilities across complex, hybrid environments. This role requires deep technical expertise in tools like Grafana, Site24x7, Azure Monitor, and a strong understanding of Azure Cloud, Windows, and Linux operating systems. The ideal candidate will also bring leadership experience, guiding teams in implementing scalable observability strategies that drive performance, reliability, and resilience.

Requirements

  • Proficiency in observability, monitoring, or site reliability engineering.
  • Expert-level proficiency in Grafana (including custom dashboards, integrations, and alerting).
  • Hands-on experience with Site24x7, Azure Monitor, and other observability platforms.
  • Strong understanding of Azure Cloud architecture, services, and deployment models.
  • Deep technical knowledge of Windows and Linux operating systems, including performance tuning and diagnostics.
  • Experience with scripting and automation (e.g., PowerShell, Bash, Python).
  • Excellent communication and team leadership skills, with a track record of mentoring and guiding technical teams.

Nice To Haves

  • Familiarity with containerized environments (Docker, Kubernetes) is a plus.
  • Certifications in Azure or related technologies.
  • Experience with OpenTelemetry, Prometheus, ELK stack, or similar tools.
  • Background in resilience engineering or incident management frameworks.

Responsibilities

  • Design and implement observability solutions using Grafana, Site24x7, Azure Monitor, and other relevant tools.
  • Develop and maintain dashboards, alerts, and metrics to monitor infrastructure, applications, and services across cloud and on-prem environments.
  • Collaborate with cross-functional teams to define SLIs/SLOs and improve system reliability and performance.
  • Lead initiatives to enhance telemetry, logging, and tracing across distributed systems.
  • Provide technical leadership and mentorship to junior engineers and cross-functional teams.
  • Troubleshoot and resolve complex issues in real-time, leveraging deep knowledge of OS internals (Windows/Linux) and cloud infrastructure.
  • Drive adoption of best practices in monitoring, incident response, and post-mortem analysis.
  • Stay current with industry trends and emerging technologies in observability and resilience engineering.

Benefits

  • Unlimited vacation, subject to local regulations and business priorities.
  • Hybrid working arrangements.
  • Paid time off for voting, bereavement, and sick leave.
  • Confidential one‑to‑one support through our Employee Assistance Program.
  • Network of Wellbeing Champions and Gather Groups.
  • Monthly events and initiatives designed to help you thrive.
  • Medical, life and disability insurance.
  • Retirement plans.
  • Lifestyle, and other benefits.
  • Paid time off for volunteering and donation‑matching opportunities.
  • Inclusion communities, such as Count Me In, Culture@Finastra, Proud@Finastra, Disabilities@Finastra, and Women@Finastra.
  • Online learning and accredited courses through our Skills & Career Navigator tool.
  • Global recognition program, Finastra Celebrates.
  • Regular employee surveys that help shape our culture and ways of working.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service