This role focuses on owning the reliability, availability, scalability, and performance of major production platforms and services. The Senior Site Reliability Engineer will work closely with Infrastructure, Development, Product, and Operations teams to ensure the ADT platform runs smoothly and customers are protected. The position involves driving operational excellence through automation, observability, and proactive problem-solving across large-scale distributed systems. Key responsibilities include leading reliability efforts in cloud environments, designing and building infrastructure as code, owning the operational lifecycle of critical services, identifying and resolving reliability gaps and performance bottlenecks, defining and improving observability practices, supporting software releases, partnering with cross-functional teams, and mentoring junior engineers. The role also requires participation in an on-call rotation for production support, incident response, and root cause analysis.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed