About The Position

We are sharing a specialised consulting opportunity for experienced Senior Backend Engineers with strong expertise in backend development, distributed systems, cloud infrastructure, DevOps, networking, IAM, observability, and production-grade systems engineering to contribute to an advanced AI training and reinforcement-learning environment project. Selected professionals will create realistic reinforcement-learning environments that test advanced AI systems on production infrastructure challenges involving system design, deployment, troubleshooting, security, scalability, and recovery. No prior experience in AI is required.

Requirements

  • Strong professional backend-engineering experience
  • Expertise in one or more of C++, Python, Rust, Go, Java, or JavaScript
  • Strong practical experience with DevOps and cloud infrastructure
  • Experience with CI/CD pipelines and infrastructure automation
  • Demonstrated ability to architect, scale, and secure distributed systems
  • Deep understanding of networking and Identity and Access Management
  • Experience with message queues, durable storage, observability, and production diagnostics
  • Strong knowledge of rolling deployments, fault tolerance, and disaster recovery
  • Experience troubleshooting production-grade systems
  • Ability to design deterministic tests and reproducible technical environments
  • Strong technical documentation and communication skills

Nice To Haves

  • No prior AI-training or model-evaluation experience is required

Responsibilities

  • Design realistic technical environments involving backend services, distributed systems, networking, queues, and durable storage
  • Create production-style scenarios covering deployment, scaling, troubleshooting, and recovery
  • Evaluate architecture for reliability, scalability, resilience, and operational correctness
  • Incorporate realistic failure modes and infrastructure constraints
  • Apply strong systems-design judgement across cloud environments
  • Develop scenarios involving IAM, authentication, authorisation, permissions, and access control
  • Work with CI/CD, infrastructure automation, rolling deployments, and operational workflows
  • Incorporate observability, monitoring, and diagnostic requirements
  • Design fault-tolerance and disaster-recovery scenarios
  • Identify common production configuration, networking, and security failures
  • Build reproducible environments with deterministic validation
  • Create golden reference solutions and objective acceptance criteria
  • Develop defective variants that test troubleshooting and recovery skills
  • Validate expected behaviour, failure conditions, and edge cases
  • Document architecture, assumptions, operational flows, and technical trade-offs clearly

Benefits

  • Independent contractor engagement
  • Fully remote
  • Output-based compensation per task
  • Minimum weekly submission requirements apply
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service