Systems Engineering, Metrics and Alerting

CloudflareSan Francisco, CA
$66,000 - $91,000Hybrid

About The Position

This role is for the internal Observability Team, responsible for the observability platform and stack to make our engineering teams productive. This includes (but is not limited to) areas like metrics, alerting, error tracking, logging, tracing, and more. In this role, you can expect to: Design, deliver, and operate software and a platform that progresses Cloudflare's Observability competency Solve scaling bottlenecks in critical services in our Metrics & Alerting pipeline Work on highly distributed and scalable systems Participate in the constant cycle of knowledge sharing and mentoring Participate in the global on-call rotation for the services your team owns Research and introduce cutting-edge technologies Contribute to open-source. We are a small team, well-funded, growing and focused on building an extraordinary company. This is a software engineering/systems engineering role and is a superb opportunity to be part of a high performing team to help to support Cloudflare’s mission and help build a better internet.

Requirements

  • A Software Engineering background and proficiency in high-level programming languages (e.g., Go)
  • Proficiency in Data structures and databases like TSDBs, Columnar stores or related
  • Proficiency in distributed Linux environments
  • Proficiency in designing high-scale distributed systems
  • Proficiency in Prometheus, Alertmanager, Thanos
  • Experience working in a fast, high-growth environment
  • Experience working in a 24/7/365 service environment
  • Exquisite written and verbal communication skills
  • Familiarity with Internetworking, networking protocols Layer 2-7 of the OSI model and BGP
  • Strong bias for action

Nice To Haves

  • Experience with high-bandwidth transit Internetworking and routing
  • Passion for code simplicity and performance

Responsibilities

  • Design, deliver, and operate software and a platform that progresses Cloudflare's Observability competency
  • Solve scaling bottlenecks in critical services in our Metrics & Alerting pipeline
  • Work on highly distributed and scalable systems
  • Participate in the constant cycle of knowledge sharing and mentoring
  • Participate in the global on-call rotation for the services your team owns
  • Research and introduce cutting-edge technologies
  • Contribute to open-source

Benefits

  • Equity plan
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service