About The Position

Ondo operates real-time trading systems that run around the clock across traditional and crypto venues. The platform spans low-latency Rust engines, a fleet of Go services for trading, execution, and PnL accounting, and a multi-region Kubernetes footprint on AWS. We are looking for an SRE with strong systems programming skills to own the reliability, observability, and performance of this platform. This is a hands-on role: you will read and modify Go and Rust code, debug latency regressions down to the feed handler, run incident response during market hours, and build the automation that keeps a 24/7 trading system healthy with a small team.

Requirements

  • 5+ years in SRE, production engineering, or infrastructure roles, with meaningful time supporting real-time or latency-sensitive systems
  • Strong programming ability in Go or Rust, and willingness to work in both; this role changes application code, not just infrastructure
  • Deep, hands-on Kubernetes and AWS experience: you have run stateful, latency-sensitive workloads in production, not just stateless web services
  • Strong observability instincts: fluent PromQL, structured-log analysis, and experience designing alerts with high signal and low noise
  • Solid Linux internals and networking fundamentals: you can chase a p99 regression through the kernel, the NIC, or the GC
  • Sound judgment under pressure and clear written communication during and after incidents

Nice To Haves

  • Experience operating trading systems, execution infrastructure, or market data infrastructure at a trading firm, exchange, or broker
  • Familiarity with market microstructure and order lifecycle (order books, order types, fills and reconciliation)
  • Experience with market data providers and protocols (Databento, SIP/prop equity feeds, venue WebSocket APIs)
  • Exposure to crypto venues and on-chain trading
  • Python for operational tooling and data analysis (pandas, parquet, BigQuery)
  • Experience with GitOps workflows, infrastructure as code, and secrets management at scale

Responsibilities

  • Debug production incidents end to end: stale market data feeds, exchange rate limits, WebSocket disconnects, order-lifecycle desyncs, and latency regressions in the trading path
  • Harden market data ingestion from providers such as Databento and venue-native feeds (REST and WebSocket), including staleness detection, failover, and replay
  • Build reconciliation and data-integrity tooling across live gauges, Postgres, and our S3 parquet data lake, so positions, fills, and PnL always agree
  • Participate in an on-call rotation covering US equity market hours and 24/7 crypto venues

Benefits

  • Competitive compensation including but not limited to salary, future token rights, and/or equity (according to your preferences)
  • Full benefits (medical, vision, and dental)
  • Flexible vacation policy (PTO)
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service