Manage team of Systems Administrators and Engineers responsible for maintaining, monitoring, patching, and remediating issues in the server infrastructure. Manage team (with lead) of database administrators responsible for maintaining, monitoring, patching, and remediating issues in our database environments. Manage relationships with 3rd party vendors including but not exclusive to DBA and Sys Admin managed services. Ensure strong incident response processes keeping all stakeholders informed as well as maintaining status pages for the organization via the Incident io product. Maintain and improve monitoring practices, looking for ways to become as close to proactive as possible. Maintain patching schedules for both OS and application related patch cycles. Optimize existing systems for performance and reliability. Develop and maintain SOPs for core functions. Lead and Support cross-functional technology projects such as upgrades, migrations, or implementations. Engage in and support DevOps practices for infrastructure and security management. Exhibit a Site Reliability Engineering mindset when engaging with or developing reliability and incident response practices. Manage projects and priorities via JIRA project management system in an Agile based format. Excellent customer service, communication, and documentation skills. Strong analytical skills with the ability to solve complex technical problems. Other duties as assigned.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior