Senior Site Reliability Engineer

BoeingBerkeley, MO
$160,650 - $217,350Onsite

About The Position

The Boeing Company is looking for a Senior Site Reliability Engineer to join the Air Dominance Site Reliability Engineering team located in Berkeley, MO. We are seeking a highly talented, motivated, and creative individual to operate, improve, and sustain mission-critical developer platforms used by Air Dominance engineering teams. This role will provide hands-on technical ownership for GitLab, GitLab CI/CD runners, Jira, Confluence, PostgreSQL, and related software delivery tools such as Artifactory and SonarQube. The selected candidate will drive reliability improvements, automate operational workflows, troubleshoot complex incidents, lead planned maintenance activities, and help establish mature Site Reliability Engineering practices for the team.

Requirements

  • Active Secret U.S. Security Clearance. (A U.S. Security Clearance that has been active in the past 24 months is considered active.)
  • Ability to obtain access to Special Access Programs (SAP)
  • Bachelor's Degree
  • 9+ years of experience with DevOps, Site Reliability Engineering, software engineering, and/or cloud engineering
  • Experience with GitLab, Azure DevOps and CI/CD (Continuous Integration and Continuous Delivery (CI/CD)
  • Experience with technical leadership
  • Experience with designing and implementing scalable computing infrastructure for data solutions, including cloud architectures (AWS, Azure, Google Cloud)

Nice To Haves

  • Bachelor of Science degree from an accredited course of study in engineering, engineering technology (includes manufacturing engineering technology), chemistry, physics, mathematics, data science, or computer science and 9+ years of related work experience or Bachelor’s Degree and 13+ years of directly related work experience or 17+ years of related, relevant experience
  • Experience administering GitLab, GitLab CI/CD, GitLab runners, or comparable enterprise source control and CI/CD platforms
  • Experience administering Jira, Confluence, or other Atlassian products
  • Experience administering PostgreSQL, including backup and recovery, replication, query troubleshooting, storage management, and performance tuning
  • Experience with AWS, Microsoft Azure, Infrastructure as Code, Ansible, configuration management, container platforms, Docker, Kubernetes, virtualization, or cloud-native services
  • Experience with Artifactory, SonarQube, Jenkins, or similar software delivery tools
  • Experience designing or improving monitoring, alerting, dashboards, operational metrics, and SLOs
  • Experience supporting Air Dominance, classified, air-gapped, or highly regulated engineering environments
  • Experience with disaster recovery planning, continuity of operations, backup validation, and restore testing
  • Experience leading technical investigations, post-incident reviews, and corrective action plans
  • Ability to obtain Security+ certification
  • Demonstrated ability to balance customer urgency, system reliability, security requirements, and change control
  • Strong customer focus and ability to work with engineering teams, leadership, and cross-functional stakeholders

Responsibilities

  • Operate and maintain GitLab, GitLab CI/CD runners, Jira, Confluence, PostgreSQL, and related developer tooling infrastructure
  • Lead the deployment and configuration of software development technologies, including build servers, version control systems, CI/CD pipelines, and automated testing frameworks
  • Lead administration of cloud-based and on-premises infrastructure using approved Amazon Web Services (AWS), Microsoft Azure, Linux, virtualization, or container platform capabilities
  • Serve as a technical owner for platform reliability, availability, performance, capacity, backup, recovery, and operational readiness
  • Lead troubleshooting for complex application, database, runner, pipeline infrastructure, network, storage, and performance issues
  • Develop and maintain Infrastructure as Code (IaC), Ansible, and other automation for provisioning, configuration, platform scaling, health checks, reporting, backup validation, and routine operational tasks
  • Plan and execute approved changes, including application upgrades, security patches, database maintenance, runner lifecycle activities, and infrastructure updates
  • Establish and improve monitoring, alerting, dashboards, SLIs, SLOs, SLAs, KPIs, error budgets, and operational metrics
  • Lead software development tool administration, maintenance, version upgrades, patch management, and integration between tools such as Jira, GitLab, Artifactory, Confluence, and SonarQube
  • Define, collect, analyze, and refine software delivery and platform reliability metrics to support data-driven decision making
  • Support incident response, root cause analysis, corrective action tracking, and post-incident reviews
  • Mentor junior engineers and provide technical guidance on SRE practices, secure administration, automation, and troubleshooting
  • Partner with developers, project administrators, cybersecurity personnel, infrastructure teams, database administrators, and program stakeholders
  • Improve runbooks, standard operating procedures, architecture documentation, and disaster recovery procedures
  • Evaluate platform risks, capacity trends, recurring incidents, and operational toil, then recommend and implement improvements
  • Lead demonstrations, monitor progress, and present technical status to customers and management
  • Lead process improvement efforts that help operationally field higher-quality end-to-end system software more frequently
  • Participate in after-hours support for urgent or mission-impacting issues as required

Benefits

  • competitive base pay
  • variable compensation opportunities
  • health insurance
  • flexible spending accounts
  • health savings accounts
  • retirement savings plans
  • life and disability insurance programs
  • paid and unpaid time away from work
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service