Data Center IT Operations Manager

Compass Datacenters•Dallas, TX
•Onsite

About The Position

This role is responsible for keeping Compass Datacenters’ internal IT environments running once facilities are live and in production. That covers the servers, networking, storage, and security infrastructure that supports our operations—not customer colocation environments. You’ll manage the team that handles day-to-day administration, monitoring, patching, and incident response across all production sites. When something breaks at 2 AM, your team is the one that gets the call. When we need to plan capacity for the next round of site launches, you’re the one building that plan. This is a hands-on leadership role—you need to be close enough to the technology to make good calls, and experienced enough as a manager to build and run a reliable team.

Requirements

  • Hands-on leadership role
  • Close enough to the technology to make good calls
  • Experienced enough as a manager to build and run a reliable team
  • Administration of Windows Server, VMware/Nutanix virtualization clusters, and associated storage environments
  • Health and performance of the production network—switches, routers, firewalls, load balancers, and WAN connectivity
  • Monitoring tools configuration and tuning
  • Active Directory, DNS, DHCP, and identity/access management
  • Backup strategy and recovery testing
  • Incident triage and escalation for security alerts
  • Patching and vulnerability management across servers, network gear, and endpoints
  • Firewall and access control management
  • Working within established security frameworks (NIST, CIS, or whatever Compass adopts)
  • Capacity planning for compute, storage, and network resources
  • Hardware and software lifecycle management
  • IT operations budget ownership
  • Site onboarding and validation
  • Direct management of systems and network administrators
  • Building and maintaining an on-call rotation
  • Vendor relationship management
  • Incident management and root cause analysis

Responsibilities

  • Oversee administration of Windows Server, VMware/Nutanix virtualization clusters, and associated storage environments across all production sites.
  • Own the health and performance of the production network—switches, routers, firewalls, load balancers, and WAN connectivity. Coordinate with the design team on changes and upgrades.
  • Make sure monitoring tools are configured, thresholds are tuned, and alerts go to the right people. Reduce noise so the team can focus on what matters.
  • Manage Active Directory, DNS, DHCP, and identity/access management across the environment.
  • Own the backup strategy and make sure restores actually work. Run periodic recovery tests—don’t wait for a real disaster to find out.
  • Serve as the escalation point for security alerts from EDR/MDR tools. Make sure threats are investigated, contained, and resolved—not just acknowledged.
  • Run the patching cycle across servers, network gear, and endpoints. Track vulnerabilities, prioritize based on actual risk, and close them out on a schedule.
  • Ensure firewall rules, WAF configurations, and access policies stay current and hardened. Clean up stale rules regularly.
  • Work within established security frameworks (NIST, CIS, or whatever Compass adopts) and support audit and compliance activities as needed.
  • Track utilization across compute, storage, and network resources. Forecast needs based on site launch schedules and growth.
  • Maintain a hardware and software lifecycle plan. Know what’s approaching end-of-life, what needs to be refreshed, and when.
  • Own the IT operations budget. Track spend against forecast, manage renewals, and make the case for capital when it’s needed.
  • Receive new sites from the construction/design team after commissioning. Validate that systems are built to standard and ready for production support.
  • Lead a team of systems and network administrators. Set clear expectations, run regular 1:1s, do real performance reviews, and invest in developing your people.
  • Build and maintain an on-call rotation that provides 24/7 coverage without burning the team out.
  • Manage relationships with hardware, software, and support vendors. Own contract renewals, escalations, and service-level accountability.
  • Run the incident response process for major outages. Lead root cause analysis after the fact and make sure corrective actions actually get implemented.

Benefits

  • Medical
  • Dental
  • Vision
  • Voluntary
  • 401K
  • Unlimited PTO for US based Employees
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service