This role involves working closely with development and operations teams to ensure efficient and reliable software deployments, system monitoring, and overall platform reliability. The Systems Engineer will be responsible for identifying and fixing system bugs, developing features using AI coding tools and script repositories for automation, scaling, testing, and securing cloud infrastructure and pipelines. Key responsibilities include enhancing performance monitoring, identifying and optimizing performance bottlenecks, contributing to the SRE journey through DevOps tools and Infrastructure-as-Code, and developing high-quality pipeline automation workflows. The role also requires developing and executing test strategies for failure scenarios, creating and running performance tests, building automated systems for continuous testing, and collaborating with SREs, developers, and operations teams to define and validate reliability goals. Ensuring new services and features are thoroughly tested before production deployment, and validating monitoring, logging, and alerting mechanisms are crucial. The position also involves tracking Service Level Indicators (SLIs) and Service Level Objectives (SLOs) and independently resolving conflicts between timeline, budget, and scope, while escalating complex issues to senior management.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level