Site Reliability Engineer Technical Lead

Bright Vision TechnologiesNew Albany, IN
$100,000 - $150,000Remote

About The Position

Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic opportunity to join an established and well-respected organization offering tremendous career growth potential. Job Summary: The Senior IT Technical Lead for Site Reliability Engineering (SRE) is a hands-on, expert-level role. This individual is responsible for ensuring the operational excellence of critical production systems. This role is less about managing project timelines and resources, and more about creating innovative solutions to improve a mission critical service. The technical lead will drive automation, optimize performance, lead incident response from a technical standpoint.

Requirements

  • At least 8 years of experience in IT, with significant time spent in a senior-level SRE or similar role.
  • Deep SRE and Systems Expertise: Comprehensive knowledge and hands-on experience applying SRE principles to manage the reliability and scalability of enterprise-level systems.
  • Cloud platforms (e.g., AWS, GCP, Azure), microservices, containers (Kubernetes).
  • Proven ability to write and implement code (e.g., Python) and leverage automation tools to eliminate manual toil and create scalable, self-healing systems.
  • Expertise in implementing robust monitoring, alerting, logging, and tracing systems (e.g., Dynatrace, Splunk, ELK Stack).
  • Able to analyze system metrics to proactively identify issues and drive data-backed technical decisions for performance tuning and capacity planning.

Nice To Haves

  • Exceptional leadership, communication, and interpersonal skills, with the ability to influence and collaborate effectively with both technical and non-technical stakeholders, including senior leadership.

Responsibilities

  • Ensuring the operational excellence of critical production systems.
  • Creating innovative solutions to improve a mission critical service.
  • Driving automation.
  • Optimizing performance.
  • Leading incident response from a technical standpoint.
  • Writing and implementing code (e.g., Python) and leverage automation tools to eliminate manual toil and create scalable, self-healing systems.
  • Building and maintaining internal tools that enhance infrastructure and operations.
  • Implementing robust monitoring, alerting, logging, and tracing systems (e.g., Dynatrace, Splunk, ELK Stack).
  • Analyzing system metrics to proactively identify issues and drive data-backed technical decisions for performance tuning and capacity planning.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service