Sr. Site Reliability Engineer (US Federal)

WorkdayReston, VA
Hybrid

About The Position

Workday, Federal Cloud Platform Engineering, is building a new SRE team responsible for deploying, operating and supporting an innovative cloud native service platform on a foundation of Kubernetes in both Public Cloud and Private Cloud environments. This provides a secure platform on which dozens of Workday service teams, and Platform development teams can build and test their prerelease code, through deployment to production on a continuous basis. This role will support one or more direct or indirect contracts with the U.S. Federal Government which, due to federal government security requirements, mandates that all Workday personnel working on the contracts be United States citizens (naturalized or native).

Requirements

  • 8+ years of SRE/DevOps experience in a distributed systems environment.
  • 8+ years' proven experience in managing and fixing distributed systems. (AWS, GCP, Kubernetes, Docker)
  • 5+ years of experience with Linux.
  • 5+ years' demonstrated ability with at least one of (GoLang, Python, Ruby), preferably GoLang (Go)
  • 5+ years of experience with Bash or Shell scripting, plus understanding of software development standard methodologies such as code management, CI/CD.
  • BS in Computer Science or related job experience
  • Ability to work independently
  • Skills and passion to operate, maintain, support and sustain the platform.
  • Excited by working in a fast-paced environment.
  • Experience collaborating with multi-functional global and remote teams with a diverse set of backgrounds.
  • Excellent documentation skills, experience with developing detailed runbooks, processes
  • Passion for identifying and solving problems on distributed environments scaling across configuration, Linux Operating System and network

Responsibilities

  • Updating the platform continuously in line with the major and minor release cycles of open source projects such as Kubernetes, Istio, Calico.
  • Providing level support for the buildout of new Customer environments
  • Collaborating with multi-functional teams to come up with automation solutions
  • Leading new Service team onboarding engagements, in partnership with Platform Engineering Architects and Product Managers
  • Overall system health, and holding engineering teams accountable to meet agreed SLO's such as latency and error rates.
  • Weekly platform release preparation, including evolving the automation towards zero touch.
  • Follow through on improvements identified post-patch.
  • Vulnerability management, holding teams accountable to meet customer facing Service Level Agreements (SLAs)

Benefits

  • Workday Bonus Plan or a role-specific commission/bonus
  • Annual refresh stock grants
  • Comprehensive benefits
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service