The Reliability Observability Engineer 3 is responsible for enabling reliable, measurable, and supportable application operations across a broad portfolio of production applications. This role helps ensure application teams have actionable visibility into service health, customer experience, dependencies, performance, and failure conditions through effective use of metrics, logs, traces, dashboards, alerts, synthetic monitoring, and service-level indicators. As a senior-level Reliability Engineer specializing in observability, this role partners closely with product owners, application engineering teams, SRE teams, and business stakeholders to translate customer journeys and business outcomes into measurable reliability objectives. The role establishes and maintains best-practice processes for documenting, governing, reviewing, and improving user journeys, SLIs, SLOs, synthetic monitoring, dashboards, alerts, telemetry standards, and related observability assets. The engineer provides senior technical guidance, identifies observability gaps through incident and performance analysis, drives continuous improvement, and helps ensure teams have the data, processes, and operating discipline needed to detect issues earlier, reduce customer impact, and improve overall service reliability.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior