The Site Reliability Engineer is responsible for ensuring the reliability, availability, performance, and recoverability of OneAZ's technology platforms and infrastructure. This position combines infrastructure engineering, automation, monitoring, and resiliency practices to maintain highly available systems supporting associates and members. The engineer designs and implements automation solutions using PowerShell and other scripting technologies, administers enterprise monitoring platforms, leads disaster recovery testing activities, and partners with technology teams to improve operational resilience, service reliability, and recovery readiness across on-premises and cloud environments. This role serves as a key contributor to incident response, infrastructure modernization, and continuous improvement initiatives focused on reducing operational risk and improving system uptime. The Site Reliability Engineer works closely with Infrastructure, Information Security, Application Support, Enterprise Architecture, and business teams to identify operational risks, strengthen recovery capabilities, and improve the overall resilience of technology services that support associates and members.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level