Cloud Infrastructure Engineer

Seamless Migration
Remote

About The Position

Seamless Migration is seeking a career and customer-oriented Cloud Infrastructure Operations Engineer to join their team. This fully remote position requires the individual to work in a highly critical production environment and serve as the first point of contact for cloud-related escalations reported by customers and application teams. The role focuses on identifying, isolating, and resolving infrastructure issues before they impact business operations, acting as an active troubleshooter using deep technical knowledge and modern AI tools. Responsibilities include executing automation scripts, taking ownership of assigned work, performing rapid triage of incoming issues, analyzing system logs, and isolating issues between application, OS, and cloud infrastructure layers. The engineer will also open and manage support cases with cloud providers like AWS and Azure, package detailed hand-offs for the L3 team when SLAs are exceeded, and leverage AI assistants for log analysis and troubleshooting. Monitoring backup status and performing restores, as well as developing small-scale automation using Bash, Python, and Ansible, are also key duties. Availability for off-hours support is required.

Requirements

  • Must be a U.S. Citizen and only hold U.S. Citizenship; no dual citizens.
  • 1–3 years of progressive IT experience, with hands-on exposure to cloud infrastructure administration and operational support.
  • Strong administration skills in Red Hat Linux.
  • Hands-on experience with AWS and Azure core services.
  • Proficiency in programming/scripting languages and automation tools such as Python, Bash, or Ansible.
  • Ability to safely execute automation scripts for cloud resource and OS-level operations, including understanding script inputs and outputs, workflow, and troubleshooting basic execution issues.
  • Familiarity with AI tools for troubleshooting, log analysis, and infrastructure support.
  • Hands-on mindset: Willingness to actively troubleshoot issues, resolve problems in real time, and communicate directly with stakeholders.
  • Sound judgment: Ability to follow established procedures, make appropriate decisions within defined guidelines, and recognize when an issue should be escalated to senior engineers.
  • Continuous learner: Eagerness to learn from senior engineers and progressively take on more complex technical responsibilities.

Nice To Haves

  • Experience using AI assistants and AIOps tools to analyze and summarize logs, identify potential issues, and accelerate troubleshooting and evidence gathering.
  • Experience developing and maintaining small-scale automation using Bash, Python, and Ansible.
  • Experience performing AWS AMI and Azure image-based VM restores.
  • Experience opening and managing support cases with AWS and Azure.
  • Willingness to provide support during off-hours, nights, or weekends when critical infrastructure issues or escalations require immediate attention.
  • Ability to communicate directly with application owners, customers, and stakeholders.
  • Eagerness to learn from senior engineers and progressively take on more complex technical responsibilities.

Responsibilities

  • Serve as the first point of contact for cloud-related escalations reported by customers and application teams.
  • Identify, isolate, and resolve infrastructure issues before they impact business operations.
  • Act as an active troubleshooter using deep technical knowledge and modern AI tools to expedite resolutions and support the senior L3 team.
  • Safely execute automation scripts for cloud resource and OS-level operations, including understanding script inputs and outputs, workflow, and troubleshooting basic execution issues.
  • Take full responsibility for assigned work, take prompt action on production issues, follow established troubleshooting and escalation procedures, and maintain a strong focus on service availability, uptime, and operational stability.
  • Act as the primary responder for incoming cloud infrastructure issues and perform rapid triage to understand the scope and business impact of reported problems.
  • Collect and analyze system logs, network traces, and filesystem information.
  • Apply Red Hat Linux skills to perform basic LVM operations, analyze syslog, and troubleshoot OS-level processes and services.
  • Contact application owners or customers directly to gather missing details.
  • Isolate issues between the application layer, OS layer, and Cloud Infrastructure layer.
  • When needed, open and manage support cases with cloud providers, including AWS and Azure.
  • If an issue exceeds defined SLAs, package all collected logs, evidence, and preliminary troubleshooting into a detailed and defined hand-off template for the L3 senior Infrastructure team.
  • Leverage AI assistants and AIOps tools to analyze and summarize logs, identify potential issues, and accelerate troubleshooting and evidence gathering.
  • Monitor AWS and Azure Backup status reports and perform VM/image-based restores such as AWS AMI and Azure image restores following established procedures.
  • Develop and maintain small-scale automation using Bash, Python, and Ansible to streamline data collection, filtering, and routine troubleshooting and evidence-gathering tasks.
  • Be available to provide support when needed during off-hours, nights, or weekends when critical infrastructure issues or escalations require immediate attention.

Benefits

  • 100% paid Medical, Dental & Vision for our employees
  • 6% 401K match (Vested Immediately)
  • 29 Days' PTO
  • Flexible Work Schedule
  • Tuition/Certification Reimbursement
  • Growth Opportunities w/in an Emerging Defense Company
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service