Lead DevOps Engineer, Vice President

State StreetToronto, ON

About The Position

This opportunity is ideal for individuals interested in contributing as a Lead DevOps Engineer within State Street’s Global Technology Services (GTS) organization. GTS is a critical enabler of State Street’s business, delivering data, insights, and digital capabilities to our clients. We are driving enterprise-wide digital transformation by leveraging industry best practices and advanced technologies. The ideal candidate demonstrates integrity, creativity, and a strong commitment to continuous learning. This role requires a collaborative mindset, a passion for innovation, and the ability to influence and inspire cross-functional teams. We are seeking a highly skilled DevOps Engineer to own, manage, automate and continuously improve the infrastructure deployment process, security, reliability, and operational excellence of our digital platform. The ideal candidate will be responsible for end-to-end DevOps operations, including cloud infrastructure management, CI/CD automation, application deployments, monitoring, security hardening, disaster recovery, performance optimization, and platform reliability. This role requires close collaboration with development, architecture, security, content management, and business teams to ensure a highly available, secure, and scalable platform.

Requirements

  • Strong experience with AWS and AWS SDK, EC2, ECS, Lambda, S3, RDS, Route 53
  • Working knowledge of Docker, Kubernetes, Amazon EKS, Linux, Bash, Python
  • Exposure of IAM, authentication and authorization layers
  • Working experience of how to deploy and manage AEM (Adobe Experience Manager) including author, publisher, dispatcher and caching layers.
  • Observability experience using CloudWatch, Splunk, Datadog, Grafana, Dynatrace
  • Knows how to use Harness, Jenkins, Jfrog
  • Knowledge of Networking, DNS, SSL/TLS, Security Hardening, Vulnerability Management, Disaster Recovery, Backup & Recovery

Nice To Haves

  • Terraform, CloudFormation, CloudFront, Elastic Cache
  • Knowledge of Incident Management, Site Reliability Engineering (SRE), Performance Optimization.

Responsibilities

  • Own the operational health, availability, performance, and security of the AEM platform and supporting AWS infrastructure.
  • Design, implement, and maintain Infrastructure as Code (IaC) solutions for cloud resources and application environments.
  • Build, maintain, and optimize CI/CD pipelines for AEM applications, frontend assets, and supporting services.
  • Manage AEM environments including Author, Publish, Dispatcher, CDN, caching layers, and integrations.
  • Implement automated deployment, rollback, and release management processes.
  • Monitor platform performance, availability, security, and user experience using modern observability tools.
  • Manage AWS services including networking, compute, storage, security, and content delivery components.
  • Implement cloud security best practices, vulnerability remediation, secrets management, and compliance controls.
  • Lead incident management, root cause analysis, and post-incident remediation activities.
  • Ensure business continuity through backup, recovery, and disaster recovery planning and testing.
  • Optimize cloud infrastructure for performance, scalability, resilience, and cost efficiency.
  • Support capacity planning, environment provisioning, and platform upgrades.
  • Collaborate with development teams to improve deployment reliability, application observability, and operational readiness.
  • Maintain platform documentation, operational runbooks, and standard operating procedures.
  • Participate in on-call support and production issue resolution as required.

Benefits

  • inclusive development opportunities
  • flexible work-life support
  • paid volunteer days
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service