AWS Cloud Platform Engineer Lead

Cynet SystemsReston, VA

About The Position

We are seeking an experienced AWS Cloud Platform Engineer Lead to oversee the design, development, and implementation of software systems and applications. This role requires strong leadership experience in driving transformation initiatives, particularly in Site Reliability Engineering (SRE) and cloud automation. You will be responsible for installing, tuning, upgrading, troubleshooting, and maintaining computer systems, developing automation techniques, evaluating new systems, and determining integration issues. Additionally, you will improve engineering job knowledge by staying current with best practices and acting as a mentor for team members.

Requirements

  • Bachelor's Degree in Information Technology or Computer Science. In lieu of a bachelor's degree, an additional 4 years of relevant work experience is required in addition to the required work experience.
  • Minimum of 10 years of IT experience of which at least 5 years must be in AWS Cloud Platform engineering and Administration.
  • Strong Leadership experience with driving Transformation initiatives.
  • 3-5 years of experience in a Site Reliability Engineering role.
  • Experience with SRE principles and transformation.
  • 3+ years of experience with Containerization (Kubernetes), Cloud technologies (AWS, Azure etc.), DevOps tool chain (Ansible, Jenkins, Artifactory, bitbucket, etc.), and technical patterns (IaC, Automated Provisioning/Release, CI/CD, etc.).
  • Solid understanding of Software coding techniques and experience with full spectrum of Software engineering (Build, Integration, Test, Releasing and Deployment) leveraging Python.
  • Experience in Developing and/or challenging engineering solutions/practices and collaborating with peers within and outside of immediate team, including customers (Dev, Architects, Engineers).
  • Minimum of One AWS certification is required.
  • Experience programming with one or more languages: Python, Java, Groovy, Go, etc.
  • Strong skills and experience in at least one IAC Tool for Platform Automation: Ansible, Terraform, AWS Cloud formation, CDK.
  • Docker or other OCI-certified containers.
  • Experience with Kubernetes, AWS EKS, AWS ECS.
  • Experience with CNI Plugins: Calico, Flannel, Weave Net.
  • Experience with Service Mesh: Istio, AWS App Mesh, OpenShift Service Mesh.
  • Experience with Container Security Tools: Twistlock, Sysdig, Aqua.
  • Experience with Platform Monitoring, Observability, & Performance Tools: Nginx, New Relic, AppDynamics, Data Dog, Thanos, Jaeger, LogDNA.
  • Experience with DevOps Tools: Git/Repo, Crucible, Bitbucket, Jira, Ansible, Puppet, Jenkins, ArgoCD, Bamboo, Maven, Artifactory, Nexus.
  • 10+ years of overall experience in IT including hands-on Development and Systems engineering background.
  • 5-10 years of experience in cloud engineering and automation, with a focus on AWS cloud services.
  • Minimum of 8-10 years of IT experience of which at least 5 years must be in AWS Cloud Automation and Administration.
  • 3-5 years of experience in a Site Reliability Engineering role.
  • 3+ years of experience with implementation of Containerization (Kubernetes), Cloud technologies (AWS, Azure, or Google, etc.), DevOps tool chain (Jenkins, Artifactory, bitbucket, etc.), and technical patterns (IaC, Automated Provisioning/Release, CI/CD, etc.).
  • In-depth knowledge of AWS services and solutions, including but not limited to EC2, S3, RDS, Lambda, VPC, IAM, CloudFormation, and CloudWatch.
  • Strong understanding of cloud architecture principles, design patterns, and best practices for building scalable, resilient, and secure cloud environments.
  • Proficiency in infrastructure as code (IaC) tools such as Terraform or AWS CloudFormation.
  • Exposure to Artificial intelligence patterns, Machine Learning and engineering of AWS solutions.

Nice To Haves

  • Master's degree preferred.

Responsibilities

  • Installs, tunes, upgrades, troubleshoots, and maintains all computer systems relevant to the supported applications including all necessary tasks to perform operating system administration, user account management, disaster recovery strategy and networking configuration.
  • Develop and implement techniques to prevent system problems, troubleshoots incidents to recover services, and support the root cause analysis.
  • Evaluates new systems by performing in-depth tests, including end-user reviews.
  • Researches software and related products to support recommendations and purchasing.
  • Determines systems integration issues by evaluating components; developing and completing performance tests; analyzing test data; studying project requirements; analyzing user and potential user input; evaluating similar and related products and systems.
  • Develop system automation and system integration of business processes.
  • Improves engineering job knowledge by attending educational workshops; reviewing professional publications; establishing personal networks; benchmarking state-of-the-art practices; participating in professional societies.
  • Acts as a mentor for junior and senior team members.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service