About The Position

Deploy, integrate, and operate high-performance storage for GPU-accelerated compute and AI platforms. You will own the storage layer where Kubernetes meets bare metal — standing up NFS-based high-performance storage, wiring it into clusters via CSI, and tuning it to keep data flowing to GPU workloads at scale. Work spans hybrid, edge, and air-gapped deployments built on the Mirantis K0rdent stack. We are looking for a senior DevOps engineer who treats storage as infrastructure to be automated, observed, and tuned — not hand-managed. The right candidate is fluent in Kubernetes storage, comfortable on bare metal down to the disk, kernel, and NFS-client layer, and knows how to make high-performance NAS actually perform under demanding workloads. You should reach for infrastructure-as-code and GitOps by default, be self-directed in diagnosing performance and reliability issues end to end, set operational standards for others to follow, and communicate clearly across teams.

Requirements

  • 5+ years in DevOps, SRE, or infrastructure operations, with strong hands-on experience operating Kubernetes storage (CSI, persistent volumes, storage classes) in production.
  • Experience integrating and operating NFS-based high-performance / NAS storage, including data-path tuning.
  • Bare-metal operations experience: host provisioning, disk/storage configuration, and Linux storage and networking fundamentals.
  • Proficiency with infrastructure-as-code (Terraform/OpenTofu) and GitOps-driven configuration.
  • Scripting/automation skills (e.g., Bash, Python, or Go).
  • Strong written and verbal communication with technical audiences.

Nice To Haves

  • Hands-on experience with VAST and/or Dell PowerScale.
  • Experience with GPUDirect Storage and RDMA/RoCE data paths.
  • Experience with the Mirantis K0rdent stack (K0rdent Enterprise, K0rdent AI, k0s, MKE) and Cluster API.
  • Familiarity with other storage backends (Ceph, object/S3) and CSI driver operations.
  • Proven experience in sovereign or high-security air-gapped environments.

Responsibilities

  • Integrate NFS-based high-performance storage (e.g., VAST, Dell PowerScale) into Kubernetes clusters via CSI, storage classes, and persistent volumes.
  • Tune the NFS data path — mount options, nconnect/RDMA, client and network settings — for high-throughput, low-latency GPU/AI workloads.
  • Deploy and operate storage services and operators; manage capacity, quotas, snapshots, and lifecycle.
  • Provision and configure storage on bare-metal hosts, including disk layout, drivers, and kernel/network tuning.
  • Deliver storage integration for k0s-based Kubernetes via Cluster API (CAPI) and K0rdent management/child cluster topologies.
  • Operate storage in fully disconnected (air-gapped) environments, including local artifact/mirror connectivity (Harbor) and PKI/TLS considerations.
  • Automate storage provisioning and configuration with infrastructure-as-code (Terraform/OpenTofu) and GitOps pipelines (ArgoCD or Flux).
  • Build monitoring, alerting, and observability for storage performance, capacity, and health.
  • Diagnose and resolve performance, reliability, and scaling issues across the storage stack.

Benefits

  • Professional development and training
  • Attend conferences and working groups
  • Company outings, happy hours, hackathons, and tech talks
  • competitive compensation package with a strong benefits plan
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service