The Site Reliability Engineer (SRE) will ensure the performance, reliability, and scalability of our systems as we continue to grow. This role bridges the gap between software development and operations, applying software engineering principles to automate, optimize, and enhance the reliability of our infrastructure and production systems. Your role includes identifying recurring failure patterns, implementing automated solutions, and continuously improving platform performance. Leveraging your intellectual curiosity and expertise in operations and development, you will also play a pivotal role in monitoring security and reliability threats, while actively advocating effective solutions. This role spans a genuinely wide range of work, from deep automation and greenfield infrastructure projects to legacy system support and cross-team collaboration. You'll thrive here if you enjoy variety and can move between priorities without missing a beat, and if you're motivated by helping shape reliability practices as we grow toward higher availability targets.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level