OSS Telemetry Observability Engineer

E-SpaceArlington, TX
Onsite

About The Position

E-Space is bridging Earth and space to enable hyper-scaled deployments of Internet of Things (IoT) solutions and services. We are building a highly-advanced low Earth orbit (LEO) space system that will fundamentally change the design, economics, manufacturing and service delivery associated with traditional satellite and terrestrial IoT systems. We’re looking for an OSS Telemetry Observability Engineer to grow our team. This role will be hands-on and you will work with our core engineering team and customer support escalation groups, to monitor and properly escalate real-time core network or infrastructure issues that could affect customers. Your role will be within a dynamic and busy Satellite communications company and will work alongside individuals who are highly experienced and experts in their field. You will maintain and improve the OSS network and provide technical support, fault investigation and troubleshooting of all issues on the network. You will keep the performance of our telecom network at optimum levels by ensuring that network/application problems are detected and corrected according to agreed KPI and SLA — and your work will directly shape the reliability and scalability of the networks our partners depend on.

Requirements

  • 10+ years of experience working in a telecom environment or similar roles
  • Expertise in monitoring tools (Prometheus, Kibana, Grafana, ELK Stack)
  • Strong knowledge of distributed systems and microservices architecture
  • Basic Python / Linux / Bash knowledge
  • Experience with cloud platforms (AWS, GCP, Azure)
  • Understanding of SLO/SLI/SLA concepts
  • Excellent written and verbal communication skills
  • Excellent interpersonal skills

Nice To Haves

  • Experience with 3GPP NTN core adaptations (Rel-17/18): satellite access node (SAN) architecture, timer and window parameter adaptation for high-latency links, or UPF placement in non-terrestrial topologies.
  • Experience in a startup or fast-paced R&D environment where architectural decisions move at product speed.

Responsibilities

  • Design and implement comprehensive observability strategies across distributed systems
  • Deploy and maintain monitoring solutions using tools like Prometheus, Kibana, ELK Stack, Grafana
  • Develop automated alerting systems with AI-powered anomaly detection
  • Create and maintain dashboards for real-time system visibility
  • Implement distributed tracing and log aggregation solutions
  • Creating on-demand dashboards to monitor specific metrics or to validate restoring of services after failures
  • Technical support to NOC engineers to analyze and solve issues.
  • Development of KPIs and other metrics to monitor key aspects in mobile network performance
  • Develop and maintain dashboards, alerts, and visualizations to track key performance metrics.
  • Development of training material related to KPIs / Dashboards
  • Ability to troubleshoot, document, and assess proper escalation channel and team or group.

Benefits

  • Competitive salaries
  • Continuous learning and development
  • Health and wellness care options
  • Financial solutions for the future
  • Optional legal services (US only)
  • Paid holidays
  • Paid time off
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service