Senior Systems Engineer (Remote)

The Home DepotTEXAS - VIRTUAL - TX01, TX
$80,000 - $150,000Remote

About The Position

The Sr. Systems Engineer is responsible for independently developing, maintaining, and supporting The Home Depot's technical infrastructure of hardware and system software that drives the success of Home Depot and our customers. As a Systems Engineer II, you will be part of a dynamic team with engineers of all experience levels who help each other build and grow technical and leadership skills while creating, deploying, and supporting production infrastructure. In addition, Sr. Systems Engineers may be involved in routine upgrades and application support as well as root cause and post-mortem analyses around security incidents and service interruptions.

Requirements

  • Must be eighteen years of age or older.
  • Must be legally permitted to work in the United States.
  • 4 years of work experience.
  • The knowledge, skills and abilities typically acquired through the completion of a bachelor's degree program or equivalent degree in a field of study related to the job.

Nice To Haves

  • Deep technical expertise in container orchestration and application platforms, specifically Red Hat OpenShift, Kubernetes, and Cloud Foundry / VMware Tanzu Application Services (TAS).
  • Drive platform automation and modern GitOps deployment workflows using Ansible, ArgoCD, and Flux to ensure consistent, secure, and declarative environment management.
  • Maintain robust observability across our platforms by implementing and fine-tuning Prometheus for metrics collection, Grafana for Observability, and Fluentbit for log stream processing.
  • Strong experience in high-availability platform architecture, infrastructure-as-code, operational troubleshooting, and enterprise supply chain service resiliency is required to maintain mission-critical uptime across hybrid environments.
  • Rotational on-call will also be required.
  • Experience working as part of a collaborative, cross-functional, modern engineering team.
  • Experience in troubleshooting and remediation within multiple Information technology disciplines.
  • Experience installing and upgrading applications or databases and performing system maintenance.
  • Familiarity with system and environment analysis, design, and optimization.
  • Familiarity with debuggers, runtime analysis, library systems, compiled programming, and software update tools.
  • Experience monitoring the operational status and performance of, and configuring as well as tuning, systems, networks, or databases.
  • Experience with operating system commands and utilities as well as scripting.
  • Experience with cloud platforms such as GCP and Azure.
  • Experience supporting a 24x7 retail operation.
  • Experience with version control systems.
  • Experience with CI/CD toolchain.
  • Experience with production system designs including Infrastructure as Code, High Availability, and Performance monitoring.
  • Exposure to Site Reliability Engineering (SRE).
  • 1+ year of previous leadership experience.

Responsibilities

  • Keeps abreast of innovations and industry trends as well as changes to internal systems and determines how they impacts tools, training, and support necessary to keep systems up, running, and secure.
  • Participates in and contributes to learning activities around modern systems engineering core practices (communities of practice).
  • Proactively views articles, tutorials, and videos to learn about new technologies and best practices being used within other technology organizations.
  • Researches and analyzes business trends and behavioral data to identify opportunities for improvements and new initiatives.
  • Drives the evaluation, development, and recommendation of specific technology to provide cost-effective solutions that meet THD requirements.
  • Researches and designs best fit infrastructure, network, database, cloud, AI, and security architectures for products.
  • Proactively creates and maintains tools for monitoring and support.
  • Participates in project planning and reporting across multiple efforts.
  • Collaborates with product and project teams to understand needs and enable them with infrastructure.
  • Supports technology architecture design review efforts for project and product teams.
  • Leverages tooling and custom applications to monitor the operational status of applications, infrastructure, networks, databases, and security; optimizes and tunes performance as appropriate.
  • Drives root cause analysis, debugging, support, and post-mortem analysis for security incidents and service interruptions.
  • Maintains, upgrades, and supports existing systems and infrastructure to ensure operational stability.
  • Opens and manages vendor problem tickets to resolution.
  • Drives the production of in-house documentation around solutions.
  • Provides application support for software running in production.
  • Drives moving KB articles to infrastructure as code models.
  • Drives keeping monitoring/alerting up to date.

Benefits

  • The pay range for this position is between $80,000.00 - $150,000.00
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service