Head of Infrastructure and DevOps

SmartcatGeorgia - Remote, GA
Remote

About The Position

Smartcat is seeking a leader to manage DevOps and Infrastructure teams. The role involves improving and maintaining the technical infrastructure to support the development, deployment, and operation of Smartcat. Key responsibilities include ensuring the availability, scalability, and reliability of staging and production environments, and defining and enforcing best SRE and DevOps practices. The goal is to achieve zero incorrect release-related production issues, maintain 99.9% availability for CI/CD and staging servers, and 99.99% product availability with effective disaster recovery plans. The role also focuses on decreasing lead time by accelerating CI/CD pipelines and environments, and building and managing a high-performing, globally distributed engineering team through strengthening the hiring process, providing technical guidance, coaching, and career development.

Requirements

  • 3 years of experience leading a team in a fast-growing multinational IT business.
  • 5 years of experience in DevOps and SRE, maintaining high-RPS (Requests Per Second) distributed products.
  • Utilize a data-driven approach with the ability to justify decisions using concrete metrics.
  • Ability to build relationships with neighboring teams, initiate and adhere to agreements, and act as an interface between your team and the environment.
  • Strong hands-on experience with DevOps practices (in a leading role), specifically in building complex CI/CD pipelines, implementing IaC (Infrastructure as Code), and realizing various deployment strategies.
  • Experience in implementing secure development environments.
  • Practical experience in developing and maintaining large scalable applications in the cloud (AWS, Azure, GCP).

Nice To Haves

  • .Net Core, MongoDB, Vue 2/3, AWS, Kafka, ES, Gitlab CI, Prometheus, Victoria Metrics, Jaeger, k8s

Responsibilities

  • Manage DevOps and Infrastructure teams to improve and maintain technical infrastructure.
  • Ensure the availability, scalability, and reliability of staging and production environments.
  • Define and enforce best SRE and DevOps practices.
  • Enforce rigorous CI/CD testing and reviews to eliminate release errors.
  • Maintain robust staging infrastructure with minimal downtime, ensuring a 99.9% uptime.
  • Implement resilient systems and procedures for near-perfect product uptime, with effective disaster recovery plans.
  • Decrease lead time by accelerating CI/CD pipelines and environments.
  • Build and manage a team of A+ players across different regions and time zones.
  • Strengthen the hiring process, provide expert technical guidance, effective coaching, and career development for team members.

Benefits

  • Remote-friendly work options
  • Global, connected team
  • Opportunity to shape the future of AI in the workplace
  • Be part of a company innovating in a $100 Billion industry
  • Join a rapidly growing company (130% YoY growth)
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service