About The Position

Become part of a team that operates enterprise-level, business-critical observability and logging platforms. Do you have in-depth experience with Elasticsearch, OpenSearch, or the ELK Stack? Are you interested in the stable operation of highly available platforms, complex troubleshooting, and the analysis of technical incidents? Then we are looking for you. As a Senior Elastic / OpenSearch Operations Engineer (m/f/d), you will be responsible for operating a business-critical logging and observability platform in an enterprise environment. Together with international teams, you will ensure that billions of log data are reliably processed, analyzed, and provided. You will work with modern technologies such as OpenSearch, Elasticsearch, Grafana, Prometheus, Kubernetes, and automation solutions, taking on a central role in platform operations. Your tasks include: Operation, monitoring, and optimization of large Elasticsearch and OpenSearch platforms, analysis and resolution of incidents in 2nd and 3rd level support, conducting root cause analyses and sustainable error correction, ensuring availability, performance, and stability of business-critical systems, monitoring cluster health and troubleshooting complex platform disruptions, optimizing index, sharding, and storage concepts, further development of monitoring and observability solutions, collaboration with international DevOps, SRE, and Platform teams, automation of recurring operational tasks, creation and maintenance of runbooks, operational documentation, and best practices, participation in a 24/7 operations organization and on-call structure.

Requirements

  • Several years of experience as: Elastic Engineer, OpenSearch Engineer, ELK Engineer, Site Reliability Engineer, Platform Operations Engineer, Observability Engineer
  • Practical experience with Elasticsearch and/or OpenSearch
  • Experience in operating productive logging or monitoring platforms
  • Knowledge in: Cluster Operations, Sharding, Replication, Index Management, Performance Tuning, Incident Management
  • Very good Linux knowledge
  • Experience with monitoring and observability solutions
  • Good English and German language skills, both written and spoken

Nice To Haves

  • Kubernetes
  • Helm
  • ArgoCD
  • Prometheus
  • Grafana
  • OpenTelemetry
  • Kafka
  • Python
  • Bash
  • Ansible
  • Infrastructure as Code
  • Cloud technologies (AWS, Azure, or GCP)

Responsibilities

  • Operation, monitoring, and optimization of large Elasticsearch and OpenSearch platforms
  • Analysis and resolution of incidents in 2nd and 3rd level support
  • Conducting root cause analyses and sustainable error correction
  • Ensuring availability, performance, and stability of business-critical systems
  • Monitoring cluster health and troubleshooting complex platform disruptions
  • Optimization of index, sharding, and storage concepts
  • Further development of monitoring and observability solutions
  • Collaboration with international DevOps, SRE, and Platform teams
  • Automation of recurring operational tasks
  • Creation and maintenance of runbooks, operational documentation, and best practices
  • Participation in a 24/7 operations organization and on-call structure

Benefits

  • Permanent employment contract
  • Attractive remuneration
  • Further training
  • Hybrid and full remote model
  • PC equipment
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service