This role requires a Site Reliability Engineer II with a strong background in site reliability engineering, DevOps, infrastructure, or production operations. The ideal candidate will have hands-on experience in incident response, observability, and automation, with the technical expertise to mentor and guide other engineers. Experience with shift-based, on-call, or follow-the-sun coverage models is essential. A working knowledge of major cloud providers, particularly AWS, and modern observability tools like Datadog, Prometheus, and Grafana is required. Proficiency in at least one scripting or programming language is necessary for automation guidance, along with a solid understanding of SLI/SLO frameworks and reliability engineering principles. Excellent English communication skills are needed for effective collaboration with international teams.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level
Education Level
No Education Listed