AI DevOps Engineer| Hybrid: Chicago, IL or NYC, NY

Circana Careers•Chicago, IL
•Hybrid

About The Position

We are seeking an experienced Senior DevOps Engineer to design, implement, automate, and support enterprise infrastructure and application delivery platforms across both on-premises and hyperscaler cloud environments. This role will be responsible for driving DevOps practices, CI/CD automation, Infrastructure as Code (IaC), platform reliability, and operational excellence in hybrid infrastructure landscapes. The ideal candidate will possess deep expertise in cloud and data center technologies, automation, container platforms, and production support. This individual will partner closely with Infrastructure, Application Development, Security, Architecture, and Operations teams to deliver scalable, secure, and highly available technology solutions.

Requirements

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
  • 7+ years of experience in DevOps, Infrastructure Engineering, Cloud Engineering, or Platform Operations.
  • Strong experience supporting both on-premises infrastructure and cloud-based environments.
  • Extensive hands-on experience with Jenkins administration, pipeline development, automation, and integration.
  • Experience with CI/CD platforms such as Jenkins and Azure DevOps
  • Experience implementing Infrastructure as Code Ansible.
  • Strong Linux administration experience.
  • Hands-on experience with Kubernetes, Docker, and container orchestration platforms.
  • Experience supporting Azure or GCP environments.
  • Strong scripting and automation skills using PowerShell, Python, Bash, or similar languages.
  • Experience with enterprise monitoring tools such as Grafana, Prometheus, or Azure Monitor.
  • Strong troubleshooting, problem-solving, and communication skills.

Nice To Haves

  • Multi-cloud experience across Azure and GCP.
  • Azure DevOps Engineer Expert, Azure Solutions Architect or related certifications.
  • Experience managing large-scale hybrid cloud transformations.
  • Knowledge of platform engineering and self-service infrastructure models.
  • Experience supporting retail, distribution, supply chain, or large-scale customer-facing platforms.

Responsibilities

  • Design, build, and maintain enterprise DevOps platforms supporting software development and infrastructure teams.
  • Develop and manage CI/CD pipelines utilizing Jenkins, Azure DevOps, or similar tools.
  • Implement Infrastructure as Code (IaC) using Ansible, or equivalent technologies.
  • Automate application deployment, infrastructure provisioning, configuration management, patching, and operational workflows.
  • Drive continuous improvement initiatives focused on deployment speed, reliability, security, and operational efficiency.
  • Support enterprise data center infrastructure including virtualized environments, compute, storage, networking, and backup solutions.
  • Manage Linux-based environments supporting business-critical applications.
  • Integrate on-premises systems with cloud platforms to support hybrid operational models.
  • Participate in infrastructure lifecycle management, upgrades, patching, disaster recovery testing, and capacity planning.
  • Deploy, automate, and support AI/ML infrastructure platforms and services in cloud and on-premises environments.
  • Support AI engineering teams with scalable environments for model training, inference, experimentation, and deployment.
  • Manage infrastructure supporting generative AI, large language models (LLMs), AI copilots, vector databases, and AI orchestration frameworks.
  • Establish operational monitoring, security controls, compliance, and reliability standards for AI solutions.
  • Design, deploy, and support workloads across Microsoft Azure and Google Cloud Platform (GCP).
  • Manage and optimize hybrid cloud environments integrating on-premises infrastructure with cloud services.
  • Support cloud governance, security standards, identity management, cost optimization, and operational best practices.
  • Partner with cloud vendors and internal teams to troubleshoot and resolve complex platform issues.
  • Develop cloud migration, modernization, and platform engineering strategies.
  • Design and support containerized application platforms using Kubernetes, OpenShift, AKS, EKS, or GKE.
  • Manage cluster administration, networking, security, scalability, and operational health.
  • Implement and maintain enterprise monitoring, logging, and observability solutions.
  • Support production systems with a focus on availability, performance, and resiliency.
  • Participate in incident management, root cause analysis, problem management, and on-call support rotations.
  • Develop operational runbooks and support procedures for critical platforms.
  • Apply Site Reliability Engineering (SRE) principles to improve system availability and reduce operational risk.
  • Ensure compliance with organizational security policies, audit requirements, and regulatory standards.
  • Support disaster recovery, business continuity, and resiliency initiatives.

Benefits

  • paid time off
  • medical/dental/vision insurance
  • 401(k)
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service