About The Position

Responsible for designing, implementing, and maintaining enterprise monitoring and observability solutions using Dynatrace. The role focuses on ensuring end-to-end visibility across infrastructure, applications, and cloud environments, improving system performance, supporting incident response, and enhancing monitoring strategies across enterprise systems.

Requirements

  • Bachelor's degree in Information Systems, Computer Science, or equivalent experience.
  • Five or more years of experience in systems engineering, infrastructure support, or monitoring roles.
  • At least two years of hands-on experience with Dynatrace or similar APM tools.
  • Experience monitoring Windows and Linux server environments.
  • Experience supporting cloud platforms such as AWS, Azure, or GCP.
  • Understanding of networking concepts such as DNS, TCP/IP, and load balancing.
  • Strong troubleshooting and analytical skills.
  • Ability to interpret system and performance metrics effectively.

Nice To Haves

  • Experience configuring and maintaining enterprise monitoring solutions.
  • Experience developing dashboards, alerts, and system health checks.
  • Experience supporting application performance monitoring and transaction tracing.
  • Experience working with virtualized environments such as VMware.
  • Experience supporting incident response and root cause analysis.
  • Experience identifying system performance trends and capacity issues.
  • Experience in regulated industries such as healthcare or financial services.
  • Familiarity with Kubernetes or container-based monitoring.
  • Knowledge of ITIL practices and processes.
  • Experience with scripting or automation using PowerShell or Bash.
  • Strong collaboration and communication skills.
  • Proactive and detail-oriented mindset.

Responsibilities

  • Configure, maintain, and optimize Dynatrace monitoring solutions.
  • Ensure end-to-end observability across infrastructure, applications, and cloud systems.
  • Develop and maintain dashboards, alerts, and system health monitoring.
  • Monitor service performance and ensure early detection of issues.
  • Define and optimize alert thresholds to reduce false positives.
  • Monitor servers, virtual environments, databases, middleware, and network components.
  • Collaborate with infrastructure and cloud teams to analyze performance trends.
  • Proactively identify and resolve potential production issues.
  • Monitor application performance, transaction flows, and user experience metrics.
  • Support synthetic and real user monitoring activities.
  • Assist in troubleshooting performance issues at system and application levels.
  • Participate in incident response and provide monitoring insights during outages.
  • Support root cause analysis and identify monitoring gaps.
  • Improve monitoring standards, documentation, and alerting strategies.
  • Integrate monitoring tools with IT service management platforms.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service