About The Position

At Cognitiv, we are redefining media buying with our Deep Learning Advertising Platform, using cutting-edge deep learning technology and data science. We are looking for a Senior Site Reliability Engineer to strengthen our AWS infrastructure and improve service management across Cognitiv. The role involves scaling and hardening our AWS environment, expanding our hybrid cloud footprint, and moving towards industry best practices. This position requires an experienced engineer who can quickly learn our environment and contribute to our long-term service management roadmap. The role partners closely with our datacenter-focused SRE, requiring working familiarity with datacenter operations to provide multi-DC coverage when needed, though deep AWS expertise is the priority.

Requirements

  • Deep knowledge of AWS infrastructure, networking, and management practices.
  • 10+ years of experience in operations, software engineering, or as an SRE.
  • Working knowledge of modern datacenter practices, with the ability to support multi-DC deployments as needed.
  • Proven experience with infrastructure as code.
  • Proficiency with Python and Bash.
  • An independent self-starter who looks at the big picture, takes ownership, and independently seeks out new challenges with creative solutions.
  • A constructive, supportive team player, with good communication and interpersonal skills.

Nice To Haves

  • AWS certifications (e.g., Solutions Architect, SysOps Administrator)
  • Experience with hybrid cloud/on-prem solutions.
  • Hands-on experience building out datacenters.
  • Willingness and interest to travel 1-2 times per quarter.

Responsibilities

  • Design, implement, and maintain infrastructure across our AWS environment, serving as the primary owner of our cloud footprint.
  • Evaluate our existing AWS architecture (compute, networking, security) and ensure we are set up for long-term scalability and growth.
  • Work across engineering and product teams to scope projects tightly to core business requirements.
  • Drive engineering-wide efforts to improve company service management around deployments, monitoring, and disaster recovery.
  • Support and help maintain our co-located datacenter deployments alongside our datacenter-focused SRE, providing coverage as needed.

Benefits

  • Medical, Dental and Vision plan for US employees & Extended Health Benefits for Canadian employees
  • 12 weeks paid parental leave + 4 weeks WFH
  • Unlimited PTO + Work-From-Anywhere August
  • Career development with clear advancement paths
  • Equity for all employees
  • Hybrid work model & daily team lunch
  • Health & wellness stipend + cell phone reimbursement
  • 401(k) & RRSP with employer match
  • Parking (CA, WA, Vancouver offices) & pre-tax commuter benefits
  • Employee Assistance Program
  • Comprehensive onboarding (Cognitiv University)
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service