Sr. Monitoring Engineer (VA ESOM)

KentroUNAVAILABLE, UNAVAILABLE
Remote

About The Position

Kentro is seeking a Senior Monitoring Engineer to support the Department of Veterans Affairs ESOM program. The Senior Monitoring Engineer is part of a team that is responsible for the design, implementation, and sustainment of enterprise wide IT monitoring capabilities for the Department of Veterans Affairs, ensuring comprehensive visibility into infrastructure, applications, and service health. In this role, the engineer develops monitoring architectures, deploys and optimizes monitoring tools, establishes alerting and reporting standards, and continually enhances monitoring coverage to support operational reliability. The position collaborates closely with operations, engineering, and program stakeholders to align monitoring strategies with mission needs and drive proactive detection, rapid incident response, and continuous service improvement.

Requirements

  • Master’s degree in computer science, electronics engineering or other engineering or technical discipline
  • 10 years of relevant Experience
  • 10 years of additional relevant experience may be substituted for education
  • Hands on experience implementing and administering Dynatace, Solarwinds or other monitoring tools, including dashboards, alerting policies, and service mapping.
  • Strong proficiency with Dynatrace and SolarWinds platform deployment, configuration, and performance tuning.
  • Experience integrating monitoring solutions with ServiceNow for automated incident, event, and CMDB workflows.
  • Deep understanding of infrastructure, application, and network monitoring concepts.
  • Experience with log aggregation, event correlation, and performance analytics tools.
  • Ability to develop dashboards, reports, and observability views for technical and leadership audiences.
  • Strong troubleshooting, analytical, and problem solving skills in complex enterprise environments.
  • Excellent communication and collaboration skills, with the ability to work across engineering, operations, and program teams.
  • US Citizen or Green card holder
  • Willing and able to obtain and maintain Public Trust Clearance
  • Must meet updated ID requirements: https://www.gsa.gov/technology/it-contract-vehicles-and-purchasing-programs/federal-credentialing-services/get-appointment-help/bring-required-documents
  • If you do not currently meet the ID requirements outlined, you must be willing and able to update your current forms of ID in a timely manner to complete the suitability process successfully.

Nice To Haves

  • Experience designing or operating observability platforms (APM, logging, metrics, tracing) at enterprise scale.
  • Certifications in Dynatrace, SolarWinds, ServiceNow, or related monitoring/ITSM technologies.
  • Familiarity with cloud platforms (AWS, Azure, GCP) and cloud native monitoring services.
  • Experience with automation and scripting (PowerShell, Python, Ansible).
  • Knowledge of container orchestration platforms (Kubernetes, OpenShift) and monitoring of microservices.
  • Experience with event correlation, AIOps platforms, or machine learning based alerting.
  • Background supporting large federal, DoD, or public sector IT environments.
  • Experience developing monitoring strategies for hybrid or modernized infrastructures.
  • Understanding of ITIL processes and how monitoring aligns with incident, problem, and change management.
  • Ability to mentor junior engineers and contribute to technical leadership within the monitoring discipline.

Responsibilities

  • Design, implement, and maintain enterprise wide IT monitoring architectures and standards.
  • Deploy, configure, and optimize monitoring tools to ensure full visibility into systems, applications, and services.
  • Develop dashboards, alerts, reports, and service health views that support proactive operations.
  • Integrate monitoring platforms with ITSM/ESM systems to enable automated ticketing and incident workflows.
  • Continuously evaluate and expand monitoring coverage to address evolving infrastructure and application needs.
  • Tune alerts and thresholds to reduce noise, eliminate false positives, and improve signal quality.
  • Collaborate with operations, engineering, and program stakeholders to align monitoring strategies with mission requirements.
  • Support incident response with monitoring insights, contributing to troubleshooting and restoral efforts.
  • Maintain documentation, runbooks, and SOPs related to monitoring tools and processes.
  • Drive automation initiatives to enhance monitoring efficiency, self service capabilities, and operational resilience.

Benefits

  • paid time off
  • healthcare benefits
  • supplemental benefits
  • 401k including an employer match
  • discount perks
  • rewards
  • education reimbursement for certifications, degrees, or professional development
  • flexibility for you to take a course, complete a certification, or other professional growth and networking
  • funds for activities – virtual and in-person – e.g., we host happy hours, holiday events, fitness & wellness events, and annual celebrations
  • charity galas/events
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service