Senior Site Reliability Engineer - Cloud Platform

GoDaddyBritish Columbia, Canada, BC
CA$107,000 - CA$161,000Remote

About The Position

Global Compute builds and operates the core cloud infrastructure that engineering teams rely on every day. We provision and manage AWS accounts across the company, operate the network backbone that connects them, and maintain the security guardrails that keep those environments safe, compliant, and scalable. We believe reliability is an engineering challenge, not an operations task. We automate repetitive work, build for scale before it becomes a problem, and invest heavily in observability to identify issues before they impact the business.

Requirements

  • 5+ years of experience in Site Reliability Engineering, Platform Engineering, Infrastructure Engineering, or a similar role supporting production AWS environments.
  • Strong Python development skills with experience building automation, services, or platform tooling.
  • Hands-on experience with Infrastructure as Code, including CloudFormation and/or AWS CDK.
  • Experience operating core AWS services including IAM, VPC networking, EC2, Lambda, and managed storage or database services.
  • Strong Linux systems administration, troubleshooting, incident response, and production operations experience.

Nice To Haves

  • Experience operating AWS environments across multiple accounts or AWS Organisations.
  • Experience with AWS networking technologies such as Transit Gateway, IPAM, or BYOIP.
  • Cloud cost optimisation or FinOps experience.
  • Experience with CI/CD pipelines and GitOps-style deployment practices.
  • Familiarity with AI-assisted engineering workflows and tooling.
  • Exposure to policy-as-code, compliance automation, or cloud governance tooling.

Responsibilities

  • Operate and scale AWS production infrastructure, owning the health of services that provision, secure, and manage accounts across GoDaddy AWS organisations.
  • Design, build, and maintain cloud platform capabilities using Python, CloudFormation, AWS CDK, and automation-first practices.
  • Drive cost optimisation initiatives that improve efficiency and deliver measurable business impact.
  • Improve observability through monitoring, alerting, dashboards, and operational tooling.
  • Participate in on-call rotations, lead incident response efforts, and drive long-term reliability improvements through blameless post-incident reviews.
  • Support strategic AWS initiatives across networking, identity, governance, and multi-account architecture.
  • Review code and designs, contribute documentation and operational runbooks, and mentor fellow engineers.
  • Leverage AI-assisted tooling to improve engineering productivity, accelerate automation, and reduce operational toil.

Benefits

  • competitive pay
  • generous time off
  • parental leave
  • healthcare
  • retirement savings program
  • health, dental, and vision insurance
  • life insurance
  • critical illness
  • AD&D
  • health care spending account
  • employee assistance program
  • paid sick time
  • paid personal time
  • paid parental leave
  • remote work options
  • paid holidays
  • paid Wellness days
  • employee stock purchase plan
  • discretionary cash bonus scheme that pays 10% of base salary based on individual and company performance
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service