Azure DevOps Engineer

RedolentSunnyvale, CA
Onsite

About The Position

This role involves developing software solutions to enable the reliability and operability of large-scale distributed systems. The engineer will build a deep understanding of system behavior, scaling, interaction, and failure points to identify risks and opportunities for remediation. Key responsibilities include implementing monitoring and reporting for production environments, building tools and automation to eliminate toil and reduce operational overhead, and creating frameworks, processes, and best practices for the engineering team. The role also involves defining meaningful Service Level Indicators (SLIs), automating critical engineering processes to minimize risk and maximize innovation speed, and managing capacity and performance to scale infrastructure on both public and private clouds. A deep dive into understanding application components to promote product scalability, stability, and performance, along with continuous delivery, performance fine-tuning, and troubleshooting, are also key aspects of this position.

Requirements

  • Minimum 2+ years on cloud technologies
  • Azure Certified resource or holding any certification on Linux OS
  • In-depth knowledge on Azure architecture
  • Hands-on experience with various CI/CD tool sets (GitHub, Jenkins, Puppet, Docker)
  • Extensive knowledge on CI/CD pipeline creation and containerization
  • Experience deploying Kubernetes ecosystem into Azure Container Service
  • Automation experience using Python/Ruby/Go – any scripting
  • The ability to partner and collaborate cross functionally across an engineering organization
  • Strong knowledge of Linux/Unix/BSD internals and experience working with open source software
  • Experience with technologies such as ZooKeeper, with a focus on reliability, automation, operability and performance
  • 5+ years of experience with programming languages (Python, Ruby, etc.)

Nice To Haves

  • Infrastructure as code a plus (e.g. Puppet, Chef, Ansible, Docker, etc)
  • Experienced with deploying web apps to cloud infrastructure (Azure, GCP.) and working with distributed, service-oriented architecture

Responsibilities

  • Develop software solutions to enable reliability and operability of large scale distributed systems
  • Build a deep understanding of how systems behave, scale, interact and fail, and use that insight to identity risks and opportunities for remediation
  • Implement monitoring and reporting of our production environments
  • Build tools and automation to eliminate toil and reduce operational overhead
  • Create frameworks, processes and best practices to be used across Engineering
  • Build meaningful, insightful and actionable SLIs
  • Automate critical portions of engineering processes, to minimize risk and maximize the speed of innovation
  • Manage capacity and performance to help scale our infrastructure both on public and private clouds around the world
  • Deep dive into learning and understanding the mechanism of every application component, and promoting product scalability, stability and performance
  • Continuous delivery, performance fine-tuning and troubleshooting
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service