As a Senior Platform Engineer on our observability platform team, you’ll design, build, and operate the observability capabilities that thousands of engineers rely on to understand the health, performance, and behavior of their applications and infrastructure. You’ll build on open telemetry standards to give teams a unified, vendor-flexible experience across metrics, events, logs, and traces—turning raw signals into actionable insight that accelerates troubleshooting, strengthens reliability, and improves the developer experience. This is a hands-on senior engineering role. You’ll engineer telemetry collection and pipelines, automate onboarding so teams get observability out-of-the-box, and deliver Observability-as-Code—monitors, dashboards, and alerts managed as versioned, reusable modules. A central goal is a true single pane of glass: you’ll correlate metrics, events, logs, and traces into one unified experience so engineers can move seamlessly from signal to root cause without switching tools. You’ll integrate with leading observability backends (for example, a SaaS APM/metrics platform such as Datadog alongside log platforms and open-source stacks like Prometheus and Grafana) while keeping the architecture standards-based and portable, so we’re never locked to a single vendor. You’ll define golden-signal and SLI/SLO standards, tune cost and cardinality, advance AIOps and anomaly detection, and coach engineering teams to raise observability maturity across the organization. You’ll also develop with AI—using AI-assisted coding tools and agentic workflows to build and refactor platform tooling faster, while keeping quality and security high.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior