Senior Manager, Cloud Operations

USPRockville, MD

About The Position

The Cloud Operations Senior Manager is responsible for the strategy, reliability, performance, and continuous optimization of AWS cloud services, data platforms, ETL workloads, and Acquia/Drupal environments that support both enterprise and customer-facing applications. This role ensures cloud and platform solutions are architected appropriately to meet business needs while maintaining security, scalability, resilience, compliance, and cost efficiency. The position leads operational delivery and service management practices, driving platform stability, automation, monitoring, incident response, and continuous improvement across critical technology services. The manager provides technical leadership and architectural guidance to cloud operations teams, ensuring alignment with enterprise standards and long-term technology strategy. Success in this role is measured by the ability to translate cloud operations and architecture into reliable, high-value business outcomes through strong stakeholder partnerships and operational excellence.

Requirements

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or related field, or equivalent combination of education and experience.
  • 5+ years of experience in cloud, infrastructure, application operations, data platform operations, or related IT operations roles.
  • 2+ years of experience leading, supervising, or mentoring technical teams.
  • AWS Cloud Operations Expertise – Deep experience managing AWS production environments, including compute, storage, networking, databases, security, monitoring, capacity planning, resiliency, backup/recovery, and cost optimization.
  • Cloud Architecture & Solution Design – Ability to evaluate business and technical requirements and recommend scalable, secure, resilient, supportable, and cost-effective AWS architectures aligned with enterprise standards.
  • AWS Data Platforms & ETL Operations – Strong knowledge of AWS data services, ETL/ELT pipelines, workflow orchestration, data warehouse operations, and data platform reliability, performance, and governance.
  • Cloud Security & Compliance – Experience implementing AWS security best practices including IAM, encryption, network security, logging, access governance, secure configuration management, and compliance controls.
  • Infrastructure-as-Code & Automation – Hands-on experience with Terraform, CloudFormation, scripting, and automation technologies to improve consistency, efficiency, and operational scalability.
  • DevOps & CI/CD Practices – Knowledge of DevOps methodologies, CI/CD pipelines, release automation, source control, and deployment processes that improve software delivery and operational readiness.
  • Monitoring, Observability & Troubleshooting – Expertise with CloudWatch, including performance analysis, alerting, dashboards, and root-cause investigation.
  • Acquia & Drupal Platform Support – Experience supporting Acquia Cloud and Drupal-based platforms, including web performance, security patching, release coordination, troubleshooting, caching, and CDN integration.
  • IT Service Management & Operational Excellence – Strong understanding of ITIL processes including Incident, Problem, Change, Request, and Knowledge Management, with experience using ServiceNow or similar ITSM platforms.
  • Cross-Functional Partnership & Technical Leadership – Ability to collaborate effectively with architects, developers, data engineers, security teams, vendors, and business stakeholders to deliver reliable, secure, and well-architected cloud service.

Responsibilities

  • Lead AWS Cloud Operations – Own the availability, performance, security, scalability, resilience, capacity management, and operational health of AWS cloud services and platforms.
  • Provide Cloud Architecture Leadership – Guide AWS platform and solution architecture decisions to ensure alignment with business requirements, enterprise standards, security policies, supportability, and cost-effectiveness.
  • Manage Data Platforms & ETL Reliability – Oversee AWS data platforms, ETL/ELT pipelines, and analytics workloads to ensure reliable, secure, and high-performing data processing services.
  • Oversee Acquia & Drupal Platform Operations – Ensure stability, performance, security, and operational readiness of Acquia-hosted Drupal applications supporting enterprise and customer-facing services.
  • Drive Reliability, Monitoring & Incident Management – Establish observability, monitoring, alerting, and incident response practices while leading root-cause analysis and continuous service improvement initiatives.
  • Advance Automation & DevOps Practices – Champion Infrastructure-as-Code, CI/CD, automation, and operational efficiency initiatives to improve deployment reliability and reduce manual effort.
  • Ensure Security, Governance & Compliance – Partner with cybersecurity and architecture teams to enforce cloud security controls, governance standards, audit readiness, risk management, and regulatory compliance.
  • Optimize Cloud Financial Management – Drive AWS cost transparency, governance, forecasting, resource optimization, and financial accountability while balancing performance, reliability, and business value.
  • Lead Service Management & Operational Excellence – Establish ITIL-aligned service management practices, maintain operational documentation, and continuously improve processes, service maturity, and user experience.
  • Develop High-Performing Teams & Partnerships – Lead cloud and platform operations teams, foster technical growth and accountability, and build strong partnerships with business stakeholders, architects, developers, vendors, and support organizations to deliver measurable business outcomes.

Benefits

  • company-paid time off
  • comprehensive healthcare options
  • retirement savings
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service