Technical Specialist

Insurance Corporation of British ColumbiaNorth Vancouver, BC
Hybrid

About The Position

The Operational Resilience team at ICBC is responsible for capabilities that support service reliability, operational visibility, event management, and disaster recovery readiness across the organization. The team manages enterprise monitoring and observability platforms and contributes to the resilience of critical business services. We have an opportunity for a Technical Specialist to join our team, with primary responsibility for administering, supporting, and enhancing ICBC’s Splunk platform. The successful candidate will also provide technical secondary support for other enterprise monitoring and observability tools, including Dynatrace, SCOM, Aternity, Omnibus, and related technologies. As a Technical Specialist within Operational Resilience, you will help improve visibility into the health and performance of critical services, support faster issue identification and resolution, and contribute to monitoring capabilities that support incident response, service reliability, and disaster recovery objectives.

Requirements

  • Strong experience supporting enterprise monitoring and observability platforms, with technical expertise in Splunk administration, cloud observability, and telemetry integration.
  • Hands-on experience administering and supporting Splunk Enterprise and/or Splunk Cloud environments, including architecture, deployment methodologies, performance tuning, and platform optimization.
  • Experience configuring and supporting Splunk for data onboarding, monitoring, alerting, troubleshooting, dashboards, reporting, and operational analytics across enterprise, cloud, and SaaS environments.
  • Experience administering, supporting observability platforms such as Dynatrace or similar, including full-stack monitoring, application performance monitoring, real user monitoring, & synthetic monitoring.
  • Technical project leadership include planning, consulting and driving outcomes to completion.
  • Collaborating and leading IS colleagues, management and staff in support of cross-workgroup projects, operational support and providing consultation.
  • Strong analytical, troubleshooting, and problem-solving skills with demonstrated experience utilizing observability platforms to identify & correlate service issues and performance concerns.
  • Excellent communication skills with both technical and non-technical audiences.

Nice To Haves

  • Linux and/or UNIX system administration.
  • Proficiency in developing scripts or automation solutions using PowerShell, Python, or similar technologies.
  • Experience with OpenTelemetry and modern observability practices, including distributed tracing, metrics collection, telemetry pipelines, and cloud-native monitoring.
  • APIs, integrations, automation frameworks.

Responsibilities

  • Acting as the primary technical specialist for ICBC's Splunk environment, including platform administration, maintenance, optimization, troubleshooting, and continuous improvement.
  • Supporting Splunk Enterprise and Splunk Cloud capabilities, including ingestion, parsing, indexing, search optimization, dashboards, reporting, alerting, and integrations.
  • Providing Level 3 technical support, advanced troubleshooting, root cause analysis, and resolution of complex platform issues.
  • Supporting and maintaining enterprise observability and monitoring platforms including Dynatrace, SCOM, Aternity, Omnibus, and related technologies.
  • Designing, implementing, and improving monitoring, alerting, and observability solutions that enhance operational awareness and service reliability.
  • Optimizing platform performance through search tuning, data lifecycle management, alert rationalization, capacity planning, and operational analytics.
  • Developing automation, integrations, and scripts to improve platform administration and operational efficiency.
  • Supporting platform upgrades, patching, implementation activities, testing, and lifecycle management.
  • Creating and maintaining technical documentation, operational procedures, standards, and knowledge-sharing materials.
  • Collaborating with technology teams across multiple domains to develop monitoring strategies, standards, and best practices.
  • Participating in major incident investigations, problem management activities, and post-incident reviews.

Benefits

  • Competitive salary
  • Comprehensive benefits
  • Collaborative work environment
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service