Reliability Engineer

AIS Cloud OneReston, OR
$138,000 - $209,000

About The Position

AIS is supporting a federal customer with a program focused on architecture and infrastructure. For this position, we are seeking a talented individual to join AIS as a Reliability Engineer. The role involves developing enterprise security architectures, frameworks, and standards; utilizing advanced forensics and integrating solutions with IT systems. The engineer will design secure architectures, manage integration projects, lead strategic initiatives, and enforce policies and standards. They will ensure integrity and scalability, develop comprehensive strategies, and optimize solutions for performance and efficiency. Additionally, the role requires leading architectural teams, building partnerships, managing knowledge, and communicating strategies and executive reports. The engineer will also provide architectural consulting, lead innovation initiatives, evaluate enterprise technologies, and build strategic partnerships.

Requirements

  • Bachelors degree in Computer Science, Information Systems, Engineering, or related field (or equivalent experience).
  • 8+ years of relevant experience supporting enterprise cloud and/or infrastructure environments.
  • Certifications: IAT-2, 1 or more cloud certifications.
  • Active Secret clearance (or higher).
  • Experience working in regulated environments and following secure engineering / documentation practices.

Nice To Haves

  • Experience supporting DoD/IC programs and mission systems.

Responsibilities

  • Responsible for the availability, performance, monitoring, and incident response, among other things, of the cloud platforms and services.
  • Ensure that everything that goes to production complies with a set of general requirements like diagrams, dependencies of other services, monitoring and logging plans, backups and possible high availability setups.
  • Manages uncaught exceptions, hardware degradation, networking problems, high usage of resources, or slow responses that could happen at any time.
  • Uses metrics such as mean time to recover (MTTR) and mean time to failure (MTTF).
  • Considered an emerging authority, who applies extensive technical expertise.
  • Develops technical solutions to complex problems.
  • Exercises considerable latitude in determining objectives and approaches to assignment.

Benefits

  • Employee Ownership
  • Continuous Learning
  • Inclusive Culture
  • Mission-Driven Work
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service