Manager of Infrastructure, Cloud Operations & IT (Chaos Conductor)

Ron Turley Associates LLCRemote, AZ
$140,000 - $160,000Remote

About The Position

RTA is seeking a Manager of Infrastructure, Cloud Operations & IT to lead the technology foundation that ensures RTA's people are productive and the Fleet360 platform is reliable, scalable, secure, and future-ready. This role encompasses Cloud Engineering, Infrastructure, DevOps, Database Engineering, Security Engineering, and IT operations, requiring a leader who can strategize at a high level while also being hands-on when needed. The ideal candidate will be curious about AI and automation's potential to improve operations, reduce toil, and enhance team performance. This is a leadership position for a technical operations leader passionate about building reliable systems, developing people, solving complex problems, and continuous improvement.

Requirements

  • 7+ years of progressive experience in cloud infrastructure, DevOps, SRE/platform engineering, infrastructure operations, or a closely related technical discipline.
  • Meaningful hands-on experience supporting AWS production environments.
  • Demonstrated experience leading and developing technical teams, including setting expectations, creating accountability, coaching performance, and helping strong technical people grow.
  • Experience owning or significantly influencing production reliability, availability, monitoring, and incident management.
  • Strong knowledge of modern DevOps, CI/CD, observability, infrastructure-as-code, and automation practices.
  • Strong understanding of modern cloud architecture and software design principles, including microservices and how infrastructure decisions affect application reliability and scalability.
  • Experience with technologies such as AWS, Docker, Terraform/CloudFormation, Grafana, Prometheus, Datadog, New Relic, Splunk, or comparable platforms.
  • The ability to translate complex technical problems into clear priorities, risks, tradeoffs, and recommendations for business and executive stakeholders.
  • Experience using AI and/or automation to improve technical, engineering, infrastructure, or operational workflows.
  • Strong organizational and prioritization skills.
  • A leadership style that combines confidence with humility. Comfortable making decisions, receiving feedback, engaging in healthy conflict, and changing direction when the facts tell you to.

Nice To Haves

  • Led across multiple disciplines such as Cloud, DevOps/SRE, Infrastructure, Security, Database Engineering, or IT.
  • Successfully led remote or distributed technical teams.
  • Implemented AIOps, intelligent observability, predictive scaling, automated remediation, AI-assisted incident response, or other AI-powered operational workflows.
  • Experience with security frameworks, vulnerability management, cloud security, or compliance environments.
  • Understanding of database resiliency, high availability, backup/recovery, and performance considerations.
  • Familiarity with AI infrastructure concepts such as model serving, compute-resource management, vector databases, or LLM integration patterns.
  • Relevant certifications such as AWS Solutions Architect, ITIL, or other cloud, DevOps, security, infrastructure, or AI-related credentials.

Responsibilities

  • Own the operational health of RTA’s cloud infrastructure and the teams and systems that support it.
  • Continuously improve reliability, availability, scalability, observability, capacity planning, incident response, database resiliency, security posture, and operational readiness.
  • Create clarity, accountability, strong operating rhythms, and professional growth across Cloud Engineering, Infrastructure, DevOps, Database Engineering, Security Engineering, and IT.
  • Develop people, establish priorities, remove roadblocks, drive accountability, and help multiple technical disciplines operate as one team.
  • Create connection, visibility, accountability, and strong communication across a distributed team.
  • Move AI and automation from experimentation to actual business and operational value.
  • Reduce manual toil, improve monitoring and observability, accelerate incident response, strengthen root-cause analysis, automate repeatable work, and increase team productivity.
  • Develop and maintain an AI and automation roadmap for Infrastructure and IT Operations, prioritizing opportunities based on business impact, reliability, efficiency, risk, and measurable ROI.
  • Guide the architecture, performance, scalability, availability, monitoring, alerting, and capacity planning of RTA’s AWS environment.
  • Strengthen CI/CD, infrastructure-as-code, deployment reliability, automation, observability, and the operational practices that help Engineering move quickly without sacrificing stability.
  • Provide calm, structured leadership during production incidents, improving response and resolution times, and ensuring post-incident reviews lead to lasting improvements.
  • Maintain a forward-looking view of RTA’s technical infrastructure, identify risks and capacity needs, and build roadmaps that support company and product growth.
  • Guide database performance, resiliency, capacity planning, backups, recovery, tuning, and high availability.
  • Guide vulnerability management, cloud security hardening, security engineering priorities, compliance initiatives, and incident-response preparedness.
  • Set strategic direction for endpoint management, identity and access management, networking, internal tooling, and the employee technology experience.
  • Evaluate AI-powered tools, AIOps, intelligent monitoring, predictive analytics and scaling, automated remediation, AI-assisted development tools, automated runbooks, and other technologies that can improve reliability and team effectiveness.
  • Stay close enough to the technology to jump in when needed for troubleshooting and support.
  • Lead, coach, develop, and create accountability across a multidisciplinary technical team.
  • Establish clear goals and operating rhythms and ensure people understand what matters and why.
  • Work closely with Engineering, Product, Support, Security, and other teams so reliability, scalability, security, and operational considerations are built into decisions early.
  • Translate complex technical topics, risks, investments, and tradeoffs into clear business language for senior leadership.

Benefits

  • 401(k) with 6% Safe Harbor match (100% vested day one)
  • Flexible PTO model built on trust and manager-approved time off
  • Cigna PPO and HSA medical plan options with company contributions ($780–$1,950 annually)
  • Garner Health HRA program: up to $1,000 individual / $2,000 family reimbursement opportunity
  • Wellness rewards up to $350 annually
  • Virtual care and mental health support resources
  • Company-sponsored benefits including dental, vision, EAP, life, STD, LTD, legal plan, and identity theft protection options
  • Additional employee perks, wellness initiatives, and discount programs
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service