This role focuses on owning enterprise observability using Dynatrace across various systems including cloud, on-prem, ERP, WMS, eCommerce, APIs, and integrations. The analyst will be responsible for designing service topology, dashboards, alerts, and health indicators that reflect business impact. A key aspect of the role involves applying SRE principles such as SLIs, SLOs, and error budgets to minimize incidents and enhance resilience. The position also aims to accelerate incident detection and root-cause analysis, lead post-incident reviews focused on systemic fixes, and identify potential reliability, performance, and capacity risks before they affect the business. Additionally, the analyst will define observability and SRE standards and empower teams to adopt them effectively.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed