Cloud Infrastructure Operations Analyst

Computer Task Group, Inc•UNAVAILABLE, UNAVAILABLE
•Remote

About The Position

CTG is seeking to fill a Cloud Infrastructure Operations Analyst position for our client. This role provides 24/7 technical support for cloud-based applications and infrastructure, including Level 1.5 incident remediation. The analyst will receive, document, analyze, and resolve incidents using established workflows and incident management processes. Responsibilities include initial troubleshooting across applications, DevOps, middleware, security, network, and cloud infrastructure environments. The role supports APIs, application services, IaaS, PaaS, SaaS, microservices, containers, Kubernetes nodes, and middleware components. The analyst will manage Application IDs and support cloud elasticity through automated resource scaling based on business requirements. They will also execute disaster recovery and manual redundancy failover procedures, monitor client environments using enterprise monitoring and application performance tools, and develop integrated service management reports. Communication of resolutions, action plans, and status updates to clients and stakeholders is crucial. The position follows ITIL-based Incident, Critical Incident, Problem, Change, and Integrated Service Level Management processes and requires collaboration with various teams to resolve complex issues.

Requirements

  • Cloud application and infrastructure operations
  • APIs, microservices, containers, and Kubernetes
  • IaaS, PaaS, and SaaS environments
  • Application, middleware, network, security, and infrastructure troubleshooting
  • ServiceNow, IBM Control Desk, and Remedy
  • Rundeck and IBM Runbook Automation
  • IBM APM, Dynatence, AppDynamics, NewRelic, Runscope, and Netcool OmniBus
  • Incident, problem, change, and service level management
  • Disaster recovery, failover, and cloud auto-scaling concepts
  • Strong analytical, troubleshooting, documentation, and communication skills
  • Experience supporting cloud-based applications and infrastructure in an enterprise environment.
  • Experience with incident management, technical troubleshooting, monitoring, and remediation.
  • Working knowledge of high-level cloud application architecture and integrated infrastructure components.
  • Experience using IT service management, monitoring, ticketing, and automation tools.
  • Excellent verbal and written English communication skills and the ability to interact professionally with a diverse group are required.

Nice To Haves

  • Experience supporting 24/7 or highly available production environments preferred.

Responsibilities

  • Provide 24/7 technical support for cloud-based applications and infrastructure, including Level 1.5 incident remediation.
  • Receive, document, analyze, and resolve incidents using established workflows and incident management processes.
  • Perform initial troubleshooting across applications, DevOps, middleware, security, network, and cloud infrastructure environments.
  • Support APIs, application services, IaaS, PaaS, SaaS, microservices, containers, Kubernetes nodes, and middleware components.
  • Manage Application IDs and support cloud elasticity through automated resource scaling based on business requirements.
  • Execute disaster recovery and manual redundancy failover procedures.
  • Monitor client environments and identify potential issues using enterprise monitoring and application performance tools.
  • Develop and provide daily, weekly, and monthly integrated service management reports.
  • Communicate resolutions, action plans, and status updates to clients and stakeholders.
  • Follow ITIL-based Incident, Critical Incident, Problem, Change, and Integrated Service Level Management processes.
  • Collaborate with application, infrastructure, security, and operations teams to troubleshoot and resolve complex issues.

Benefits

  • competitive benefit package
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service