Senior Software Engineer - Site Reliability

Funded.clubToronto, ON
CA$140,000 - CA$180,000Remote

About The Position

Windscribe is seeking a Senior Software Engineer to join their Engineering team. This is a software engineering role focused on writing software that eliminates operational work, rather than performing operational tasks manually. The engineer will design, write, review, test, and ship software to enhance the availability, scalability, latency, and efficiency of services. This includes automating repeated tasks, defining and enforcing SLOs with monitoring and alerting, debugging production issues across the entire stack, leading blameless postmortems, planning capacity and cost with data, and participating in the critical incident response team through an on-call rotation. The role involves learning the company's stack in depth, including DNS at global scale, anycast networks, and VPN protocols.

Requirements

  • Bachelor's degree in Software Engineering or Computer Science or similar.
  • 5+ years of professional software development in one or more general-purpose languages. Go preferred; Python, Rust, and C/C++ are also great. Shell and YAML alone do not qualify.
  • Automation you personally built — a tool, service, operator, distributed system, or auto-remediation system that eliminated real operational work.
  • Experience designing, analyzing, and troubleshooting distributed systems in production.
  • Git, testing, code review, and CI as your normal workflow; comfortable in a large codebase you didn't write.
  • Solid Linux fundamentals, containers, and infrastructure as code (Terraform, Ansible, or equivalents).
  • Observability as a builder: you instrument your own code (Prometheus, Grafana, or equivalents), not just read dashboards.

Nice To Haves

  • Deep DNS knowledge (BIND / PowerDNS / Unbound, record types, resolution paths, DNSSEC)
  • Routing experience: unicast, anycast, BGP
  • Linux networking depth: iptables / nftables, eBPF, performance tuning
  • VPN protocols (OpenVPN / WireGuard / IKEv2)
  • Experience with bare metal fleets, hypervisors, or networking hardware (JunOS / VyOS)
  • High-availability databases (MySQL, Postgres, Redis) and load balancers (HAProxy, nginx)

Responsibilities

  • Design, write, review, test, and ship software that improves the availability, scalability, latency, and efficiency of our services.
  • Turn operational work into engineering work: every repeated task, checklist, or validation becomes automation.
  • Define SLOs with stakeholders and build the monitoring and alerting that enforces them.
  • Debug production issues across the whole stack, from application code down through Linux and the network.
  • Lead blameless postmortems and root cause analyses, then land the fix that eliminates the entire class of problem.
  • Plan capacity and cost with data; document what you build so others can operate it.
  • Join our critical incident response team in an established on-call rotation.
  • Learn our stack in depth — DNS at global scale, anycast networks, VPN protocols — with the team that runs it.

Benefits

  • Stock options
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service