About The Position

Our engineering team builds next-generation, cloud-native bare-metal orchestration platforms for AI neo-cloud environments. We operate on a T-shaped engineering model: while each engineer brings deep expertise in a primary domain, every team member maintains practical fluency across neighboring systems. This shared foundation enables rigorous design reviews, effective cross-domain code reviews, and dependable on-call coverage. Our development is fundamentally based on open-source projects, and contributing to them is a major part of our work. We develop primary control plane software in Rust and Go, leverage Kubernetes operators, and foster an OpenTelemetry-first operational culture. We are seeking a Senior Software Engineer who combines strong software architecture principles with deep technical domain expertise in Datacenter Networking and DPU architectures. Fitting into our T-shaped engineering model, you will drive the development of control planes powering high-density, high-throughput infrastructure for AI training and inference workloads. Familiarity with automated bare-metal provisioning and lifecycle management is a strong asset.

Requirements

  • Primary mastery of Rust (Tokio async runtime, Tonic, Axum, sqlx) and strong proficiency in Go.
  • Advanced SQL / PostgreSQL fluency, gRPC/Protobuf contract design, and schema evolution.
  • Strong background in async concurrency models, lock-free patterns, distributed state handling, and mock-driven testing discipline.
  • Expertise in BGP, MP-BGP, EVPN, VXLAN, L3VNI, route targets, and route server design.
  • Hands-on experience with SONiC, Cumulus Linux, and NVUE (featuring a first-class NVUE client).
  • Deep knowledge of NVIDIA DOCA, Host-Based Networking (HBN), BlueField DPU architectures, and the DPF (DOCA Platform Framework) operator model (BFB, DPUSet, DPUNode CRDs).
  • Understanding of InfiniBand fabrics, NVLink / NMX-M partitioning, and RoCEv2.
  • Advanced grasp of Linux netlink, network namespaces, routing tables, and internal DHCP/DNS service implementations (the project ships its own services).
  • Familiarity with CNI plugins (Calico, Cilium).
  • Expertise in PXE/iPXE, UEFI, Secure Boot, measured boot, and TPM attestation.
  • Proficiency in Redfish, IPMI, and BMC abstractions across heterogeneous hardware platforms (Dell, Lenovo, NVIDIA reference hardware).
  • In-depth knowledge of boot chains, systemd, initramfs, BIOS configuration matrices.
  • Practical experience with PKI, X.509 certificates, TLS, SPIFFE/SVID, Vault, KMS, Keycloak (OAuth2/JWT), and RBAC.
  • Experience with KubeVirt / KVM.
  • Experience writing custom controllers/operators using kube-rs or controller-runtime with CRD-driven reconciliation loop patterns.
  • Commitment to test-driven design, writing highly testable code with mock interfaces for external hardware and network dependencies.
  • Proven track record of writing crisp architectural specs and engaging in constructive, cross-functional design reviews.
  • Bachelor's or Master's degree in Computer Science, Computer Engineering, or equivalent practical experience.
  • 5+ years of software engineering experience in cloud-native platforms, systems programming, networking, or infrastructure automation.

Nice To Haves

  • OVS and DPDK experience is a plus.
  • Demonstrated participation or maintainership in open-source systems projects is a plus.

Responsibilities

  • Design, build, and maintain production-grade control plane microservices, custom Kubernetes operators, and robust reconciliation engines in Rust and Go.
  • Author clear Functional (FR) and Non-Functional Requirements (NFR), architectural specs, system sequence diagrams, and clean gRPC/Protobuf and REST API schemas.
  • Develop custom integration modules and integrations for modern network operating systems (SONiC, NVUE, Cumulus) and DPU hardware platforms.
  • Instrument services end-to-end using OpenTelemetry (OTel) traces and metrics, performing systematic root-cause analysis across polyglot distributed systems.
  • Participate in rigorous, review-gated pull request workflows across Rust, Go, and SQL codebases while ensuring high test coverage via mock-driven testing.

Benefits

  • Professional development and training
  • Attend conferences and working groups
  • Company outings, happy hours, hackathons, and tech talks
  • Receive a competitive compensation package with a strong benefits plan
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service