About The Position

MissionOne Media, a brand within BarkleyOKRP, is seeking a Senior Infrastructure Engineer to serve as the operational backbone of their engineering team. This role initially involves taking operational ownership of the legacy environment while also contributing to the architecture and strategy for future projects. The engineer will ensure the stability, security, and high availability of core production VMs running Java/Tomcat applications and MySQL 5.7 instances, managing incidents, diagnosing root causes, tuning performance, and maintaining these critical legacy systems. Concurrently, the role supports long-term growth by partnering in the development of new products and services on modern tech stacks, balancing legacy reliability with the design and scaling of next-generation cloud architecture.

Requirements

  • 5+ years in a System Administration, Site Reliability Engineering (SRE), or Infrastructure Operations role within an enterprise environment.
  • Deep expertise in Linux administration, performance tuning, and troubleshooting (including storage and file system management, specifically ZFS).
  • Proven experience running, tuning, and debugging Java/Tomcat applications in production.
  • Strong working knowledge of MySQL 5.7, including query tuning, replication, backup/restore strategies, and migration troubleshooting.
  • Proficiency in shell scripting and writing lightweight automation (Bash).
  • Practical experience with Google Cloud Platform and managed services, including operational databases (Cloud SQL) and data warehousing platforms (BigQuery).
  • Production experience architecting and running containerized environments using Docker and Kubernetes to support new product stacks.
  • Proficiency with Infrastructure-as-Code tools (Terraform) and modern CI/CD deployment workflows.
  • Calm under pressure: A methodical, safety-first approach to live-site incident response.
  • Pragmatic problem solver: Understands the realities of legacy software and prefers practical, stable solutions.
  • Strong communication skills to collaborate with developers and stakeholders during high-stakes migrations or outages.

Nice To Haves

  • Familiarity with Java development, enabling you to read application code, assist developers with debugging, or understand internal stack traces more deeply.
  • Experience with transitioning legacy tech stacks to modern GCP / cloud-native setups.

Responsibilities

  • Own the incident management lifecycle for production, including triaging alerts, troubleshooting OS-level bottlenecks, debugging performance issues, and driving rapid root-cause analysis.
  • Actively monitor system performance, application logs, and resource health using modern observability platforms to quickly diagnose and resolve operational issues.
  • Proactively maintain, patch, and secure existing virtual machines on GCP.
  • Monitor MySQL 5.7 instances, manage backups, and troubleshoot connection pool exhaustion or locking issues.
  • Collaborate to design, provision, and scale forward-looking cloud infrastructure for new products and services, incorporating modern cloud-native tools and managed services.
  • Gradually replace manual operational tasks with automation scripts and infrastructure tools to reduce toil and improve repeatability.
  • Ensure systems comply with enterprise security baselines through rigorous OS hardening, access controls, and vulnerability patching.

Benefits

  • multiple health insurance options
  • flexible PTO
  • life insurance
  • 401K
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service