About The Position

We are seeking a highly skilled Windows Server Automation Engineer with expertise in VMware to join our team for a long-term contract. The ideal candidate will possess SME-level knowledge in Windows automation, logging, and administration. This role involves building an integrated monitoring framework, developing automated alerting and proactive response mechanisms, aggregating and normalizing data from diverse sources, and creating a unified dashboard for real-time insights across our complex infrastructure. You will leverage a variety of tools and scripting languages to deliver scalable automation solutions.

Requirements

  • Proficiency in Python, Ansible, PowerShell, and shell scripting (Bash/Korn).
  • Ability to develop automation workflows for monitoring, alerting, and remediation.
  • Hands-on experience with Splunk, Dynatrace, and other enterprise monitoring platforms.
  • Familiarity with log aggregation and parsing from multiple sources (OS, applications, infrastructure components).
  • Strong understanding of Linux (RHEL) and Windows Server environments.
  • Exposure to VMware, IBM Power/AIX, and IBM LinuxOne systems.
  • Knowledge of storage arrays, SAN switches, network switches, and IP traffic monitoring.
  • Experience with backup platforms (Rubrik, Data Domain, Infinibox).
  • Familiarity with database systems (Oracle, SQL Server, MySQL, MongoDB).
  • Ability to build dashboard interfaces for real-time monitoring and alerting (using frameworks like Flask/Django for Python or similar).
  • Ability to aggregate and normalize data from multiple sources for unified alerting.
  • Understanding security logs and compliance requirements for infrastructure monitoring.
  • Ability to identify gaps in current monitoring and design innovative solutions.
  • 5+ years in infrastructure automation or systems engineering roles.
  • Proven track record in building automation frameworks and monitoring solutions.
  • Experience working in large-scale, distributed environments with global teams.

Nice To Haves

  • Prior involvement in proactive alerting and automated remediation projects is highly desirable.

Responsibilities

  • Build an integrated monitoring solution leveraging existing tools (Splunk, Dynatrace, security platforms) and local logs from Windows/Linux servers, infrastructure components (storage arrays, SAN switches, network devices), databases (Oracle, SQL Server, MySQL, MongoDB), backup systems (Rubrik, Data Domain, Infinibox), compute nodes (Dell servers), VMware environments, IBM Power/AIX, and IBM LinuxOne.
  • Develop intelligent alerting mechanisms and automated remediation workflows to reduce manual intervention and accelerate incident resolution.
  • Aggregate and normalize data from multiple sources, including platform tools and local logs, to fill visibility gaps and provide actionable insights.
  • Create a common GUI-based dashboard for real-time monitoring, alerting, and reporting across all infrastructure layers.
  • Utilize Ansible, Python, PowerShell, shell scripting, and GUI development to deliver scalable automation solutions.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service