Observability Engineer (Splunk)

SS&CJacksonville, FL
Hybrid

About The Position

As a leading financial services and healthcare technology company based on revenue, SS&C is headquartered in Windsor, Connecticut, and has 27,000+ employees in 35 countries. Some 20,000 financial services and healthcare organizations, from the world's largest companies to small and mid-market firms, rely on SS&C for expertise, scale, and technology. SS&C combines proprietary technology with deep industry expertise to support complex financial and health care operations. Our teams design, implement, and operate solutions that help clients manage data, automate processes, and scale their businesses with confidence. You will work with industry experts, modern platforms, and evolving technologies, gaining exposure to real-world operational challenges and large-scale enterprise environments.

Requirements

  • Degree in Computer Science or related degree
  • 5+ years of experience
  • Splunk Enterprise Certified Admin (or equivalent experience).
  • SmartStore configuration and operational experience.
  • AWS fundamentals (EC2, EBS, VPC) and hybrid connectivity patterns.
  • Telemetry pipelines from containerized applications; OpenTelemetry familiarity.
  • OpenShift/Kubernetes exposure; configuration management tools (Salt/Ansible); containers/K3S.
  • Ability and willingness to teach and uplift team practices.

Nice To Haves

  • Curiosity and continuous learning.
  • Strong judgment and clear thinking under pressure.
  • Ownership mentality and bias toward automation and repeatability.
  • Communication that is direct, respectful, and documented.

Responsibilities

  • Assist in the day-to-day health, reliability, and performance of Splunk Enterprise (on‑prem), including installs, upgrades, monitoring, backups, and recovery.
  • Design, operate, and evolve distributed Splunk architectures, including Search Head Clusters (SHC) and Indexer Clusters.
  • Manage ingestion and indexing pipelines end-to-end, including forwarders, parsing rules, and index design/retention.
  • Develop and maintain SPL searches, dashboards, alerts, and data models that are operationally meaningful and actionable.
  • Operate storage and retention strategies; contribute to SmartStore planning/implementation where appropriate.
  • Partner with infrastructure and application teams to onboard telemetry correctly and sustainably (logs, metrics, traces, and events).
  • Define and enforce observability standards (naming, tagging, retention, alert thresholds, and dashboard patterns).
  • Improve signal-to-noise: reduce alert fatigue and increase actionable, well-contextualized alerts.
  • Execute production changes independently and safely: instance resizing/replacement, storage migrations, network/IP cutovers, and maintenance-window execution.
  • Apply sound engineering judgment to minimize downtime and preserve rollback options.
  • Create and improve runbooks, automation, and post-change verification checklists.
  • Plan changes with impact analysis, communications, validation, and rollback.
  • Demonstrate production rigor: backups, change windows, and post-change verification.
  • Maintain clear documentation and handoffs; leave systems better than you found them.

Benefits

  • medical, dental, and vision coverage
  • a 401(k) plan with company match
  • paid time off, holidays, and parental leave
  • professional development reimbursement opportunity
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service