Manager, Server Operations - Linux

HerbalifeTorrance, CA
Hybrid

About The Position

Herbalife is seeking a hands-on Manager, Server Operations – Linux to lead a dedicated team of systems engineers and help modernize the infrastructure that supports our global business. In this player-coach role, you will guide day-to-day Linux operations, strengthen reliability and security, and help advance our hybrid cloud strategy across on-premises data centers and leading cloud platforms such as AWS, Azure, and GCP. This is a chance to create a visible impact in a global infrastructure environment while staying close to the technology. You will have the ability to improve how Linux operations are delivered, expand automation, modernize legacy processes, and help shape engineering standards that strengthen system reliability and business continuity. If you enjoy leading people, solving complex technical problems, and building scalable operational practices, this role offers the right mix of leadership, ownership, and hands-on engineering influence. At Herbalife, you will be part of a global organization where technology plays an important role in supporting teams, customers, and business operations around the world. We offer an environment where expert infrastructure leaders can take ownership, collaborate with dedicated colleagues, and help build scalable, secure, and reliable platforms for the future.

Requirements

  • 7+ years of Linux systems administration, engineering, or operations experience in enterprise environments.
  • 2+ years of experience leading, supervising, mentoring, or managing technical team members.
  • Strong hands-on experience with on-premises infrastructure, including bare metal and virtualization, as well as cloud operations in AWS, Azure, and/or GCP.
  • Experience supporting hybrid-cloud workloads, server migrations, infrastructure modernization, and enterprise-scale Linux environments.
  • A demonstrable ability to use automation and infrastructure-as-code tools such as Ansible, Terraform, Puppet, Chef, Bash, or Python to improve reliability and reduce manual effort.
  • Experience operating in a 24x7 enterprise environment with incident management, change control, on-call programs, and SLA accountability.

Nice To Haves

  • Deep knowledge of Linux distributions such as RHEL/CentOS, Ubuntu, or equivalent, including kernel tuning, storage, networking, package management, and security hardening.
  • Automation and scripting experience with Ansible, Terraform, Bash, and/or Python; CI/CD pipeline or GitOps experience is a plus.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, Nagios, or equivalent platforms.
  • Familiarity with ITSM platforms such as ServiceNow for incident, change, and problem management.
  • Relevant certifications are a plus, including RHCE/RHCSA, AWS/Azure/GCP Associate or Professional, Linux Foundation Certified System Administrator, or ITIL Foundation.
  • Good communication and documentation skills, with the ability to partner across teams and maintain clear runbooks and operational standards.

Responsibilities

  • Own the reliability, performance, capacity, and security of enterprise Linux environments across global data centers and cloud platforms.
  • Lead, mentor, and develop a small team of Linux systems engineers, crafting clarity around priorities, workload, on-call coverage, and professional growth.
  • Drive automation and infrastructure-as-code practices using tools such as Ansible, Terraform, Puppet/Chef, Bash, and Python to improve consistency, speed, and operational efficiency.
  • Lead all aspects of patching, vulnerability remediation, OS lifecycle management, and Linux hardening standards to support secure and compliant operations.
  • Lead fixing and handling blocking issues for sophisticated incidents, including root-cause analysis, post-incident reviews, and long-term corrective actions.
  • Partner with cloud engineering, networking, storage, security, and application teams to support hybrid-cloud workloads, migrations, and infrastructure modernization.
  • Improve monitoring, alerting, observability, runbooks, and on-call practices using platforms such as Prometheus, Grafana, Datadog, Nagios, or similar tools.
  • Contribute to infrastructure roadmap planning, capacity forecasting, vendor discussions, licensing, and budget planning in partnership with technology leadership.

Benefits

  • Group Health Programs
  • Medical
  • Dental
  • Vision
  • Health Savings Account (HSA)
  • Flexible Spending Accounts (FSA)
  • Basic Life/AD&D
  • Short-Term and Long-Term Disability
  • Employee Assistance Program (EAP)
  • 401(k) plan
  • Wellness Incentive Program
  • Employee Stock Purchase Plan (ESPP)
  • Supplemental Life/Critical Illness/Hospitalization/Accident Insurance
  • Pet Insurance
  • Company-observed U.S. Holidays
  • Floating Holidays
  • Vacation
  • Sick Time
  • Volunteer Program
  • Paid Maternity and Paternity Leave
  • Bereavement Leave
  • Personal Leave
  • time off for voting
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service