About The Position

You will join our platform engineering team as a Technical Leader, helping lead the design and development of large-scale data plane systems that power the Splunk data ingestion infrastructure. The team builds high-throughput, fault-tolerant distributed systems that ingest and process observability data, including metrics, logs, and traces, at scale. You will help set technical direction, mentor engineers across multiple teams, and influence architecture decisions from prototype through production.

Requirements

  • Bachelor’s degree or higher in Computer Science, Electrical Engineering, or a related technical field with 10+ years of experience building and operating large-scale distributed systems in production, with strong proficiency in programming languages such as Golang.
  • Experience with modern observability platforms, designing pipelines for logs, metrics, and traces for storage and processing on different backend systems such as Splunk, Datadog, Prometheus, Grafana, or similar technologies.
  • Experience designing observability for large systems, not just consuming it.
  • Experience with one major cloud provider (AWS, GCP, or Azure) at a platform level, including networking, storage, and cost optimization.
  • Experience in full-lifecycle agentic development, with a good understanding of context and harness engineering best practices and a drive to explore the constantly evolving space.
  • Technical leadership experience and cross-functional influence, including experience operating at the Staff or Principal level: driving multi-team technical decisions, producing architecture proposals, mentoring senior engineers, and aligning technical strategy with business goals without requiring formal authority.

Nice To Haves

  • Proven track record owning systems that process high-volume data streams, with demonstrated skill in capacity planning, traffic management, SLI/SLO definitions, and incident response at scale.
  • C++ proficiency.
  • Practical experience with container technologies, including Docker/OCI containers and running them at scale in Kubernetes.
  • Experience designing systems with security-first principles.
  • A strong operational mindset: designing for failure, instrumenting systems, participating directly in on-call responsibilities, and owning services/components through operations.
  • Strong written and verbal communication skills.
  • Ability to recognize when to prototype quickly and when to slow down for rigor.

Responsibilities

  • Architect and lead the development of high-throughput, fault-tolerant distributed systems that ingest and process observability data (metrics, logs, traces) at scale.
  • Define and drive the technical roadmap for platform reliability, scalability, and operational efficiency.
  • Partner with product and engineering leadership to translate business requirements into technically sound, pragmatic designs.
  • Establish best practices and standards for observability and capacity planning, participate in incident response, and lead blameless post-incident reviews.
  • Mentor and level up senior engineers; raise the overall technical bar through design reviews and code reviews.
  • Own systems end-to-end — from design through deployment, monitoring, and post-incident analysis.

Benefits

  • medical, dental and vision insurance
  • a 401(k) plan with a Cisco matching contribution
  • paid parental leave
  • short and long-term disability coverage
  • basic life insurance
  • grants of Cisco restricted stock units
  • 10 paid holidays per full calendar year
  • plus 1 floating holiday for non-exempt employees
  • 1 paid day off for employee’s birthday
  • paid year-end holiday shutdown
  • 4 paid days off for personal wellness
  • 16 days of paid vacation time per full calendar year, accrued at rate of 4.92 hours per pay period for full-time employees (non-exempt)
  • flexible vacation time off program (exempt)
  • 80 hours of sick time off provided on hire date and each January 1st thereafter
  • up to 80 hours of unused sick time carried forward from one calendar year to the next
  • Additional paid time away may be requested to deal with critical or emergency issues for family members
  • Optional 10 paid days per full calendar year to volunteer
  • annual bonuses
  • performance-based incentive pay
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service