Cloud System Administrator 2

Wyetech•Laurel, MD
•Onsite

About The Position

Wyetech is seeking a Cloud Systems Administrator Level 2 to support a mission-driven cyber platform built with Java and Free and Open-Source Software technologies, including Kubernetes, Hadoop, and Accumulo. The selected candidate will help sustain and optimize a cloud-based enterprise platform that enables execution of data-intensive analytics on managed infrastructure. As a Cloud Systems Administrator, you will be part of the Operations Team responsible for maintaining day-to-day platform stability, supporting customer operations, and troubleshooting technical issues across complex Linux-based environments. This is an on-call position providing Tier 1 through Tier 3 operational support. The ideal candidate thrives in a fast-paced technical environment where attention to detail, proactive task management, and the ability to independently diagnose and resolve operational issues are essential to mission success.

Requirements

  • At least three (3) years of experience performing system administration and monitoring of large distributed systems consisting of multiple clusters, clustering across at least three racks of equipment, and a minimum of 60 nodes per site.
  • Seven (7) years of experience demonstrating a strong understanding and working knowledge of core Linux operating-system components.
  • Experience managing user and group accounts in LDAP.
  • Experience configuring and administering DHCP, DNS, and TFTP.
  • Five (5) years of experience writing software scripts using Bash, Perl, or Python.
  • Experience diagnosing and troubleshooting large-scale cloud-computing systems.
  • Experience with distributed systems used for storage and retrieval of data, such as Hadoop, Cassandra, Scality, Swift, Gluster, Lustre, GPFS, Amazon S3, or comparable big-data/high-performance computing technologies.
  • Experience with configuration-management technologies including Puppet and Salt.
  • Expert understanding of end-to-end Linux PXE/network provisioning, including Anaconda Kickstart configurations, RAID-controller utilities, TFTP images, and disk-detection scripts.
  • Experience remotely accessing and troubleshooting systems to perform hardware diagnosis and repair using technologies such as VNC, Serial over LAN, IPMI, and BIOS-level configuration.
  • Understanding of enterprise/corporate architecture and familiarity with OpenSSL and Java keystore manipulation.
  • Experience troubleshooting commodity hardware platforms.
  • Advanced knowledge of SSH tunneling and protocols, including implementation of dynamic SOCKS proxies.
  • Familiarity with SSH-based utilities such as rsync, pdsh, pdcp, and WinSCP.
  • Understanding of low-level networking concepts, including VLANs, port-channel bonding, and Layer 2/Layer 3 switch interactions.
  • Familiarity with software load balancers used in large-scale web-service environments, including HAProxy and NGINX.
  • Ability to troubleshoot network and infrastructure issues affecting distributed cloud platforms.
  • Experience with Kubernetes orchestration services and Docker images.
  • Experience with log aggregation, monitoring, and search technologies, including Elasticsearch, Logstash, Filebeat, Grafana, and rsyslog.
  • Experience monitoring system health and identifying operational or performance issues across distributed environments.
  • Demonstrated ability to work within a predefined, mission-focused team structure and receive technical guidance from senior technical resources.
  • Demonstrated willingness to learn new technologies and leverage senior resources to expand technical capabilities.
  • Ability to work independently on complex technical tasks.
  • Willingness and ability to educate and train junior technical personnel.
  • Ability to plan, communicate, lead, and oversee complex technical activities involving multiple groups.
  • Ability to operate effectively in a fast-paced, mission-critical environment.
  • Ability to support after-hours/on-call operational requirements.
  • Minimum three (3) years of applicable experience is required.
  • Candidates must also satisfy the position-specific Linux, scripting, and distributed-systems experience requirements outlined above.
  • Active TS/SCI security clearance with current polygraph is required.
  • Due to federal contract requirements, United States Citizenship and position appropriate security clearance is required.

Nice To Haves

  • A degree in Engineering, Systems Engineering, Computer Science, Mathematics, or a related technical discipline is highly desired and may be considered equivalent to two (2) years of experience.
  • Hadoop/Cloud System Administrator Certification or comparable Cloud System/Service Certification is required.

Responsibilities

  • Support daily operations of cloud-based analytic platforms.
  • Provide Tier 1–3 operational support for infrastructure and system-level issues.
  • Implement, troubleshoot, administer, monitor, and maintain large distributed computing clusters.
  • Diagnose and troubleshoot large-scale cloud-computing systems and distributed storage environments.
  • Troubleshoot operational issues within complex Linux/UNIX environments.
  • Manage Red Hat/CentOS provisioning using Kickstart.
  • Perform Linux/UNIX networking analysis and troubleshooting using tools such as nmap and tcpdump.
  • Support DNS administration, including hosts files and DHCP.
  • Assist with containerized and distributed-platform technologies.
  • Collaborate with engineering teams to identify and resolve system-performance and reliability issues.
  • Support on-call operational requirements in a mission-focused environment.
  • Work independently on complex technical tasks while collaborating within a predefined mission-focused team structure.
  • Plan, communicate, lead, and oversee complex technical tasks requiring interaction across multiple technical groups.
  • Assist with educating and training junior technical resources.

Benefits

  • The company automatically contributes 20% of each employee's gross compensation to a Simplified Employee Pension (SEP) IRA, with no requirement for employee matching.
  • All contributions are fully vested from day one, ensuring immediate ownership of retirement funds.
  • Wyetech provides a generous PTO plan of up to 200 hours annually, aligned with applicable state leave regulations.
  • Employees have the flexibility to adjust their PTO allocation at the start of each calendar year, ensuring it meets their evolving needs.
  • A Choice of Medical Plan Options, some with Health Savings Account (HSA)
  • Vision and Dental
  • Life and AD&D Benefits
  • Short and Long-Term Disability
  • Hospital Indemnity, Accident, and Critical Illness Insurances
  • Optional Identity Theft and Legal Protection Services
  • Employee Referral Bonus Eligibility up to $10,000
  • Mobility Among Wyetech-supported Contracts
  • Various contract and work locations throughout Maryland, Virginia, Colorado, Texas, Utah, Alaska, Hawaii and OCONUS
  • Various team-building events throughout the year such as: monthly lunches, summer company picnic, and an annual holiday party.
  • Employees receive two complimentary branded clothing orders annually.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service