Software Engineer, Observability

LyftToronto, ON
CA$108,000 - CA$135,000Hybrid

About The Position

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. This role focuses on maintaining, improving, and developing tooling and systems that enhance the reliability, scalability, and efficiency of Lyft's platform. The engineer will collaborate with cross-functional teams to improve observability, define and monitor service-level objectives (SLOs), and automate repetitive tasks. Participation in on-call rotations and incident response is also a key aspect of the position.

Requirements

  • 3+ years of experience working on teams responsible for software development, automation, and systems engineering.
  • Bachelor's Degree or equivalent experience in Computer Science or a relevant discipline.
  • Proficiency in creating production-ready code in one or more high-level languages, such as Go or Python.
  • Experience operating infrastructure in public cloud environments, such as AWS, including familiarity with Managed Services.
  • Experience in building and maintaining observability infrastructure to support robust monitoring and analysis.
  • Familiarity with Kubernetes and managing multi-cluster environments in production settings.
  • Proven track record with modern Observability stack, including proficiency in Prometheus, Grafana, Loki, and other open-source tracing and alerting frameworks.

Responsibilities

  • Maintain, improve, and develop tooling and systems that enhance the reliability, scalability, and efficiency of our platform.
  • Assist engineering teams in defining service-level objectives (SLOs) and provide the necessary tooling to monitor and balance feature development speed and reliability.
  • Maintain and analyze metrics from operating systems, control planes, and applications to assist in fault detection and performance enhancement.
  • Collaborate with cross-functional engineering teams to enhance Lyft's observability and meet developers' needs, ensuring alignment with design and production readiness reviews, platform management, and capacity planning.
  • Keep and maintain our documentation at a world-class level by documenting infrastructure operations processes and insights.
  • Identifying repeatable actions, and automating repetitive tasks.
  • Participate in our team's on-call rotations, respond to incidents, and support other teams to mitigate customer-impacting events.

Benefits

  • Extended health and dental coverage options, along with life insurance and disability benefits
  • Mental health benefits
  • Family building benefits
  • Child care and pet benefits
  • Access to a Lyft funded Health Care Savings Account
  • RRSP plan with company match to help save for your future
  • Flexible paid time off policy for salaried team members
  • 15 days paid time off for hourly team members, with an additional day for each year of service
  • 18 weeks of paid time off for new parents (biological, adoptive, and foster)
  • Subsidized commuter benefits and Lyft ride credits
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service