The Sr. Systems Engineer will be responsible for analyzing, architecting, designing, configuring, installing, and managing enterprise server, storage, and hyperconverged infrastructure. This role involves end-to-end administration of the Nutanix environment, including AHV clusters, Prism Central and Prism Element, node and cluster expansion, AOS and LCM upgrades, Nutanix Files and Objects, and capacity planning. The engineer will maintain up-to-date knowledge of server, storage, and hypervisor vendor technologies and associated software, and engage in technical administration of physical and virtual servers running Windows Server and Linux, along with clusters, load balancers, and out-of-band management. A key aspect of this role is designing and maintaining automation for provisioning, patching, configuration, and remediation using Ansible Automation Platform, PowerShell, and Python, treating manual, repeatable work as a defect to be automated. The position also requires writing, reviewing, and maintaining infrastructure as code and playbooks in source control, configuring, monitoring, and troubleshooting enterprise systems equipment and the supporting monitoring and alerting stack, and resolving complex issues spanning compute, storage, network, OS, and application layers. The engineer will assist other teams with server installations and configurations, provide automated build standards for self-service, and work with vendors such as Nutanix, Microsoft, and enterprise storage and backup providers. Designing, testing, and documenting backup, replication, and disaster recovery, including restore testing and recovery time validation, is also a core responsibility. This role includes mentoring other Systems Engineers, raising the team's automation and scripting skill level, providing on-call support, overseeing troubleshooting activities, performing system administration functions for continuous operations, and performing system infrastructure expansion, relocation, and consolidation. Developing and maintaining systems documentation, runbooks, and architecture diagrams, supporting patching, hardening, and audit/compliance requirements through documented change control, and providing root cause analysis for outages and delivering SOPs are also key duties.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior