AVP, Monitoring Engineer

LPL FinancialFort Mill, SC
$125,145 - $208,575Onsite

About The Position

Build a career that matches all your initiative with an impressive dose of innovation. From cutting-edge resources and a collaborative environment to the freedom to make an impact and more, you’ll find the ingredients you need at LPL Financial to shape your success while helping clients pursue their financial goals. Observability at LPL only works if a senior engineer is willing to live in the dashboards, the noisy alerts, and the post-incident reviews. As AVP, Monitoring, you'll be that engineer — hands-on across CloudWatch and X-Ray, Dynatrace, Grafana / Prometheus, OpenSearch (with the legacy ELK stack where it still exists), and ServiceNow — raising the bar on SLOs and alert quality without owning a team. If you'd rather pair on a noisy alert chain than collect status, this is your seat. Job Overview: As the AVP, Monitoring Engineer, you are a hands-on senior cloud observability engineer in the Monitoring pod within the Foundations team in LPL's Cloud Center of Excellence (CCOE). You partner with the VP, Monitoring and with every other CCOE team and pod — the peer Foundations pods (Security & Governance, FinOps, Functional Design Engineering & Strategy, Network Engineering), plus the Platforms, Containers, Support, and Delivery teams — to raise the quality of LPL's observability across the multi-account landing zone. The stack spans AWS-native services (CloudWatch, X-Ray, OpenSearch), Grafana / Prometheus (including Amazon Managed Prometheus and Managed Grafana where appropriate), Dynatrace, the legacy ELK stack where it still exists, and ServiceNow for ITSM and incident ticketing. You partner closely with the Support team on incident response and with the Functional Design Engineering & Strategy pod on the observability paved road for application teams. LPL is an AWS-first CCOE: a multi-account landing zone with 100+ private reusable Terraform modules that enable 60+ AWS services, all delivered through Terraform Cloud and GitHub Actions. You spend the majority of your time hands-on in dashboards, alerts, Terraform, and incident response across LPL's US offices and India Global Capability Center (GCC) — your impact comes from technical depth, code review, and peer mentorship rather than positional authority.

Requirements

  • 7+ years of progressive technical experience including 3+ years in a senior cloud security, network security, or cloud infrastructure engineering role; Bachelor's degree in Computer Science, Engineering, or a related discipline (or equivalent work experience)
  • 3+ years of hands-on production AWS at scale in a multi-account landing zone with strong production Terraform delivered through Terraform Cloud and GitHub Actions
  • 3+ years experience operating as a senior individual contributor (AVP, Senior Engineer, Staff Engineer, or equivalent), influencing technical direction and uplifting peer engineers without direct authority — including code review leadership, design-review participation, and technical mentorship
  • 3+ years experience personally participating in 24x7 production on-call rotations in a fast-paced, security-conscious, regulated environment (financial services strongly preferred)
  • 3+ years experience tuning SLOs, alerts, and dashboards across AWS-native services (CloudWatch, X-Ray, OpenSearch), Grafana / Prometheus, Dynatrace (or comparable APM), and ServiceNow ITSM integration in a fast-paced, regulated environment

Nice To Haves

  • Master's degree in Computer Science, Engineering, or MBA
  • Experience building, scaling, or leading globally distributed engineering teams across the US and India / GCC
  • Experience integrating agentic AI / GenAI tooling (Cursor, Claude Code, Copilot, Bedrock, MCP) into platform, IaC, and engineering practice
  • Strong scripting / programming proficiency in Python, Bash, or PowerShell
  • AWS Solutions Architect - Professional
  • AWS Certified Generative AI Developer - Associate
  • HashiCorp Certified: Terraform Associate (004) or Authoring & Operations
  • Certified Kubernetes Application Developer (CKAD)
  • Open-source contributions, public technical writing, or conference speaking on cloud, IaC, or platform engineering topics
  • Experience with Backstage or another Internal Developer Platform (IDP)
  • Experience with FinOps practices and cloud cost management at scale

Responsibilities

  • Hands-on author and curate the CCOE observability paved road: opinionated Terraform modules, Helm charts, and reference dashboards for CloudWatch, Grafana, Dynatrace, and OpenSearch — so every workload starts with credible observability
  • Raise the bar on SLOs, golden signals, and alert quality across the multi-account landing zone — kill noisy alerts, surface missing ones, and partner with the Support team on what should and should not page a human
  • Operate and continuously improve the CCOE observability stack: CloudWatch and X-Ray, Grafana / Prometheus (including Amazon Managed Prometheus and Managed Grafana), Dynatrace, OpenSearch (and the legacy ELK stack where it still exists), and ServiceNow for ITSM and incident ticketing
  • Partner with the Support team on incident response: alert routing into ServiceNow, runbook execution, major incident participation, and the feedback loop from incident review back into durable monitoring improvement
  • Mentor Engineer 2 and Senior Engineers in the Monitoring pod through code review, design partnership, and pairing on noisy alert chains — uplift the pod without direct reports
  • Embed agentic AI capabilities into the team's engineering practice (e.g., Cursor, Claude Code, Bedrock, MCP servers, agentic IaC and review workflows) and into the platform's self-service experience for internal customers
  • Use agentic AI capabilities in day-to-day observability work: AI-assisted alert triage and noise reduction, dashboard authoring from natural-language intent, on-call copilots that summarize signal during incidents, and MCP-backed agents over telemetry
  • Operate as a hands-on senior cloud engineer: spend the majority of your time in Terraform code, security tooling configuration, vulnerability remediation, design reviews, peer reviews, and incident response — hands-on engineering is the primary leverage point
  • Personally participate in 24x7 on-call rotations as a senior technical responder and escalation point for production incidents
  • Partner with peer engineers, AVPs, and VPs across the Cloud Center of Excellence — the five CCOE teams (Foundations, Platforms, Containers, Support, Delivery) and the five Foundations pods (Security & Governance, FinOps, Functional Design Engineering & Strategy, Network Engineering, Monitoring) — to align roadmaps and remove cross-team and cross-pod blockers
  • Champion AWS Well-Architected Framework adoption (with emphasis on the Security pillar) and drive continuous improvement against operational, security, reliability, and compliance outcomes
  • Contribute to the private Terraform module library and the Account Factory for Terraform (AFT) foundational base layer, including security-control modules and reference patterns
  • Raise engineering quality across the pod through code review, design partnership, and technical pairing — acting as a force multiplier without direct reports
  • Participate in Agile/Scrum ceremonies (sprint planning, standups, backlog grooming, retrospectives) and partner with the RTE and PMO on delivery commitments and dependencies
  • Represent the pod's security posture in architecture review boards, internal audit, and customer engagements; communicate technical risk and trade-offs clearly to engineers and to non-technical executives

Benefits

  • 401K matching
  • health benefits
  • employee stock options
  • paid time off
  • volunteer time off
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service