This role ensures the resilience and optimal performance of enterprise container management platforms across multi-cloud and hybrid infrastructures. It primarily focuses on resolving complex technical challenges within Kubernetes-as-a-Service environments to maintain high availability and operational stability for customers. This role independently makes critical technical decisions to diagnose and resolve complex, high-impact incidents within production container environments. Decisions directly influence system uptime and customer satisfaction, often requiring rapid problem-solving under pressure. The incumbent regularly engages with enterprise customers to provide technical guidance and incident updates. Collaboration extends internally to product engineering teams, SRE specialists, and fellow support engineers to resolve deep-seated platform issues. This role directly impacts the operational continuity and reliability of mission-critical enterprise container platforms, with failures potentially incurring significant business costs and reputational damage. The scope encompasses complex multi cloud and hybrid infrastructure environments, ensuring systemic availability across a fleet of applications. A highly analytical and methodical working style is essential for deep technical problem-solving and root cause analysis. The role requires exceptional adaptability and resilience to navigate high-pressure situations and complex, evolving cloud-native technologies.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level
Education Level
Associate degree