About The Position

The Associate Technical Services Specialist is responsible for ensuring the quality, reliability, security, and operational readiness of cloud-hosted applications and supporting cloud environments. This role combines cloud engineering and administration, quality engineering, test automation, DevOps practices, and application reliability to improve how software is developed, validated, deployed, and operated. This individual partners with application developers and technical teams to manage cloud environments, automate testing and deployments, establish quality controls, improve observability, and proactively identify issues before they impact production.

Requirements

  • Experience with AWS, Azure, or GCP; administering and troubleshooting cloud-hosted applications, services, networking, IAM, secrets, and cloud security.
  • Experience with CI/CD pipelines, deployment automation, Git-based development, containers, and infrastructure-as-code technologies such as Terraform.
  • Experience developing automated API, UI, integration, regression, performance, and resiliency testing; demonstrating a reduction in bugs and improvements in application performance.
  • Programming or scripting experience with Python, PowerShell, Bash, JavaScript, or similar technologies, with knowledge of APIs, databases, integrations, and cloud architectures.
  • Experience with monitoring, logging, metrics, alerting, and troubleshooting using Grafana, Splunk, Datadog, or similar technologies.
  • Experience supporting highly available or business-critical applications, including incident troubleshooting, reliability, and operational readiness.
  • Strong analytical and problem-solving skills with the ability to diagnose issues across applications, infrastructure, integrations, and cloud environments.

Nice To Haves

  • Cloud Certification: AWS, Azure, or GCP certification preferred.

Responsibilities

  • Administer & Support Cloud Environments across development, testing, and production, ensuring reliable application hosting and operations.
  • Provision, Configure & Troubleshoot Cloud Resources, including compute, storage, connectivity, access, and application configurations.
  • Monitor Cloud Performance & Reliability across availability, capacity, performance, and overall infrastructure health.
  • Automate Cloud Operations & Infrastructure Management using infrastructure-as-code, configuration management, and automation to improve consistency and efficiency.
  • Design & Maintain Automated Testing Capabilities for cloud-hosted applications across development, testing, and production environments.
  • Develop Comprehensive Test Automation covering APIs, integrations, regression, smoke, UI, and end-to-end application workflows.
  • Build Reusable Testing Frameworks & Increase Coverage across critical application functionality, integrations, and cloud environments.
  • Improve Quality & Reduce Manual Testing by identifying recurring defects, automating regression prevention, and expanding repeatable test execution.
  • Integrate Automated Testing & Quality Controls into CI/CD pipelines to improve release consistency and application quality.
  • Establish Automated Quality Gates & Validation through pre-deployment and post-deployment testing for production releases.
  • Develop Reusable Deployment & Recovery Patterns across applications, including automated rollback and recovery capabilities.
  • Troubleshoot Pipeline & Deployment Failures across application, configuration, infrastructure, and cloud environments.
  • Implement Application & Cloud Observability through health checks, synthetic monitoring, logging, metrics, tracing, dashboards, and alerting.
  • Monitor Critical Services & Dependencies to proactively identify reliability risks, performance issues, and single points of failure.
  • Support Incident Analysis & Prevention through root-cause analysis and preventative controls addressing production incidents and recurring issues.
  • Automate Remediation & Recovery to improve application resiliency, reduce manual intervention, and accelerate service restoration.
  • Establish Production-Readiness & Release-Quality Criteria to ensure applications meet defined standards before deployment.
  • Support Production Deployments & Release Validation to confirm successful implementation and operational readiness.
  • Ensure Operational Resiliency Before Release by verifying monitoring, alerting, rollback, and recovery capabilities.
  • Standardize & Automate Operational Processes to reduce manual effort and improve consistency across cloud, DevOps, and quality engineering.
  • Develop Reusable Engineering Standards & Self-Service Capabilities that enable scalable, efficient, and repeatable development practices.
  • Drive Metrics-Based Continuous Improvement & Shared Ownership across development teams for quality, reliability, and operational readiness.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service