About The Position

The Active Directory Monitoring & Observability Specialist is responsible for the design, implementation, administration, and continuous improvement of enterprise monitoring, observability, alerting, analytics, and operational intelligence capabilities supporting Active Directory and hybrid identity platforms. This role partners closely with Active Directory Engineering, Identity & Access Management (IAM), Cyber Security, Infrastructure Operations, Site Reliability Engineering (SRE), and Platform Engineering teams to ensure the availability, performance, security, compliance, and resiliency of critical identity services.

Requirements

  • Hybrid schedule requirement of a minimum of 3 days per week
  • Deep expertise in Active Directory, including Forest and Domain Architecture, Domain Controller Operations, AD Replication, LDAP, DNS, Kerberos, Group Policy Management, and Security
  • Strong experience with Grafana dashboard development, metrics visualization, alerting framework design, and data source integration
  • Proficiency with Splunk, including Splunk Enterprise, Search Processing Language (SPL), Data Models, and dashboard development
  • Experience with additional technical skills such as Prometheus, OpenTelemetry, Elasticsearch, Kibana, Azure Monitor, Microsoft Sentinel, Dynatrace, AppDynamics, DataDog, ServiceNow, PowerShell, Python, REST APIs, Terraform, Git, and CI/CD Pipelines

Responsibilities

  • Design, develop, and maintain enterprise monitoring solutions for Active Directory Domain Services (AD DS), Domain Controllers, Global Catalog Servers, DNS Services, Kerberos Authentication Services, LDAP Services, Group Policy Infrastructure, Active Directory Replication, Active Directory Certificate Services (ADCS), and Microsoft Entra ID / Hybrid Identity integrations
  • Develop proactive health monitoring for replication latency and failures, authentication failures, Kerberos ticket anomalies, LDAP performance degradation, and other critical events
  • Implement real-time service health monitoring and enterprise operational dashboards
  • Build automated service health scoring and operational readiness reporting
  • Design and develop advanced Grafana dashboards and visualizations, creating executive, operational, engineering, and security-oriented dashboards
  • Integrate Grafana with platforms such as Splunk, Prometheus, Elasticsearch, OpenTelemetry, Azure Monitor, and Microsoft Sentinel
  • Build service maps, dependency views, SLA/SLO scorecards, capacity analysis dashboards, and trend analysis reporting
  • Establish enterprise observability standards and dashboard governance
  • Implement observability best practices aligned with SRE principles, defining and managing SLIs, SLOs, and KPIs
  • Develop automated anomaly detection capabilities and build predictive operational analytics
  • Facilitate root cause analysis, trend analysis, and support incident and problem management processes

Benefits

  • Behavioral Health Platform
  • Medical, Dental, Vision
  • Health Savings Account
  • Voluntary Hospital Indemnity (Critical Illness & Accident)
  • Voluntary Term Life Insurance
  • 401K
  • Sick Pay (for applicable states/municipalities)
  • Commuter Benefits (Dallas, NYC, SF, and Illinois)
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service