About The Position

CAS is seeking a highly motivated DevOps Engineer to join our Technology organization. In this role, you will help design, build, automate, and support the infrastructure and deployment processes that power our products and services. You will work closely with software engineers, architects, security teams, and product stakeholders to enhance platform reliability, scalability, performance, and security. The ideal candidate is passionate about automation, cloud technologies, continuous improvement, and enabling development teams to deliver high-quality solutions efficiently.

Requirements

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or related field, or equivalent experience.
  • 5-8+ years of experience in DevOps, Site Reliability Engineering, Cloud Engineering, or related technical roles.
  • Experience with cloud platforms such as AWS, Azure, or Google Cloud.
  • Experience building and maintaining CI/CD pipelines using tools such as Azure DevOps, GitHub Actions, Jenkins, or GitLab.
  • Deep hands-on experience running production Kubernetes clusters at scale (EKS specifically, given your AWS environment)
  • Strong grasp of workload scheduling, resource management, autoscaling (HPA/VPA/Cluster Autoscaler or Karpenter)
  • Experience with stateful workloads — critical for you since MarkLogic clusters and Spark jobs have different persistence/networking needs than typical stateless microservices
  • Custom operators, CRDs, or Helm chart authorship
  • Strong Terraform experience
  • Module design, state management (remote state, locking), and workspace strategies for multi-environment setups
  • EKS internals (control plane, node groups, Fargate vs. EC2 nodes)
  • IAM/RBAC integration (IRSA), VPC networking, security groups
  • ArgoCD or Flux experience
  • Pipeline design (GitHub Actions, GitLab CI, Jenkins, or CodePipeline)
  • Prometheus/Grafana, OpenTelemetry
  • Log aggregation (Loki, CloudWatch, ELK)

Nice To Haves

  • Experience with EMR-on-EKS or running Spark on Kubernetes is a strong plus

Responsibilities

  • Design, implement, and maintain cloud-based infrastructure and platform services.
  • Build and manage CI/CD pipelines that support automated testing, deployment, and release management.
  • Develop and maintain Infrastructure as Code (IaC) solutions using industry-standard tools.
  • Monitor system health, performance, availability, and capacity across production and non-production environments.
  • Collaborate with software engineering teams to improve application deployment, observability, and operational excellence.
  • Implement and maintain security best practices throughout the software delivery lifecycle.
  • Troubleshoot and resolve infrastructure, deployment, and application-related issues.
  • Support disaster recovery, backup, and business continuity initiatives.
  • Identify opportunities for automation and process improvements to increase efficiency and reliability.
  • Participate in on-call support and incident response activities as required.
  • Can write and defend architecture decision records (ADRs), technical documentation, standards, and operational procedures.
  • Cost optimization mindset (right-sizing, spot instances, reserved capacity)
  • Security-first thinking (pod security standards, network policies, secrets management via Vault/External Secrets).

Benefits

  • generous vacation plan
  • medical, dental, vision insurance plans
  • employee savings and retirement plans
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service