This role focuses on Site Reliability Engineering (SRE) principles to ensure the scalability, stability, and performance of systems. The SRE will be responsible for gathering and analyzing metrics, participating in system design and capacity planning, and working closely with the incident response team to restore services. A key aspect of the role involves balancing feature development speed with reliability and service-level objectives, while also investigating and mitigating unwanted traffic. The SRE will establish continuous process improvement cycles and partner with development teams to enhance services through testing and release procedures.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed