As a DevOps Engineer III, you will be part of the L3 support team for Operations across Edge/on-prem and cloud, owning complex incidents end-to-end: triage, deep-dive debugging, root-cause analysis, remediation, and follow-ups. Strong Linux administration (RHEL primarily, plus Ubuntu) and OpenShift/Kubernetes expertise are essential. To reduce Operations (Customer Deployment) issues, you will build targeted automations (Python, Bash, Ansible) and automate new and existing SOPs used by Operations. You will execute safe deployments and upgrades via GitOps and IaC pipelines (Flux, Ansible, Terraform) on AKS and GKE—coordinating validation and rollback plans—and contribute to the maintenance of existing GitLab CI/CD pipelines together with the DevOps engineering teams. You will design and continuously refine Alertmanager rules and standardize actionable Grafana dashboards with Operations, ensuring effective use of Prometheus metrics and logs (Grafana Alloy, Thanos). Beyond day-to-day operations, you’ll apply deep DevOps, CI/CD, and infrastructure automation expertise, drive best practices, share knowledge through workshops and mentoring, write and maintain documentation and SOPs (Standard Operating Procedure), test infrastructure, and collaborate across teams to optimize systems and workflows.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed