Senior Software Engineer, Reliability

Tin CanSeattle, WA
$175,000 - $195,000

About The Position

Tin Can is building a safer, simpler way for kids to connect — without smartphones. We’re creating screen-free, delightful devices and services that let families call the people who matter most, free from the noise of today’s digital world. We’re building a bold, authentic, nostalgic, and kinda quirky brand that resonates with folks who want something simpler & better for their kids than the tech-infused lives we’re currently living (and who have a sense of humor about it). As we gear up to scale to thousands of families, we’re ready to bring on an engineer to help make it happen.

Requirements

  • 5+ years building and operating production backend or distributed systems where downtime is something real people notice
  • Real depth in AWS and infrastructure as code (we use Terraform), and a bias toward reproducibility
  • You’ve owned reliability as a discipline, not a fire drill — SLOs, meaningful alerting, load testing, and incident review that changes what gets built next
  • Comfortable in real-time and telephony systems, or genuinely eager to get there; SIP, RTP, and media behavior are the substrate here
  • Fluency using AI as a multiplier across diagnosis, analysis, and testing
  • Hands-on and scrappy; you’d rather run the load test or read the SIP trace yourself than wait on someone else
  • A clear communicator who writes things down
  • A collaborative, low-ego approach — everyone here sweeps the floor

Nice To Haves

  • production FreeSWITCH, Kamailio, or carrier SIP trunking
  • load-testing a system through a seasonal peak
  • Postgres at scale
  • operating a consumer hardware fleet
  • observability tooling like Grafana

Responsibilities

  • Own the reliability outcome families actually experience — did the call connect, did it sound right, was the phone reachable — and the number we hold ourselves to
  • Make change management real for customer-visible behavior, so a migration that changes what a family hears is announced rather than discovered
  • Own the pipelines that carry an application or call-path change into production: reviewed, tested, and reversible
  • Own the staging tier as a real gate, where a change proves itself against production-like traffic before a kid’s phone sees it
  • Load-test the activation chain ahead of our holiday peak, so the busiest morning of the year is one we’ve already rehearsed
  • Build the product-outcome layer of our observability — the funnels and SLOs showing whether a call, voicemail, or activation actually succeeded — alongside the infrastructure metrics our SRE owns
  • Drive incidents to real root cause across the stack, from SIP signaling and media to Lambda services and Postgres
  • Own the incident-review practice: every class of failure gets an owner, a runbook, and a test that would catch it next time
  • Partner with our reliability PM to turn diagnosis into a backlog engineering can actually execute against

Benefits

  • We’re building tech that protects childhood
  • You’ll own the promise the whole company is judged on
  • Small, high-trust team
  • Room to explore, not just execute
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service