Site Reliability Engineer- FedRamp

Cisco•Rtp, NC
•Hybrid

About The Position

The Webex Contact Center FedRAMP team builds and operates secure, reliable cloud services for U.S. government customers. We partner closely with application engineering, security, compliance, and infrastructure teams to support the Webex for Government environment and deliver services that meet strict availability and control requirements. The team is made up of cloud operations, platform, and security-minded engineers who thrive in a highly regulated, high-visibility environment. This role will deploy, operate, and maintain resilient AWS and Kubernetes-based infrastructure and microservices that support the Webex for Government environment. You will monitor service health, troubleshoot complex incidents, and help restore service quickly in a 24x7 production environment. You will automate repeatable operational work and infrastructure provisioning to improve safety, efficiency, and consistency across deployments. You will strengthen platform security and compliance by supporting identity and access management, logging, hardening, vulnerability remediation, and FedRAMP continuous monitoring activities. You will collaborate across engineering, security, and compliance teams to identify operational risks, improve reliability, and build documentation, runbooks, and incident reports that make the environment easier to operate.

Requirements

  • Bachelor’s degree in Computer Science, engineering, or related field, plus 5+ years of related experience, or equivalent practical experience.
  • 5+ years of Software development or automation experience using Java, Go, Python, or a comparable programming language, including testing, troubleshooting, and maintaining code.
  • Experience supporting cloud infrastructure, production services, DevOps, platform engineering, or Site Reliability Engineering in AWS.
  • Experience with Linux administration and troubleshooting application, system, or networking issues in production or production-like environments.
  • Experience with Kubernetes, Docker, microservices, and cloud-native application deployment/troubleshooting, including Git, CI/CD pipelines, and infrastructure-as-code/automation (e.g., Terraform, Python, Bash, or Go).

Nice To Haves

  • Experience operating services in a FedRAMP, government cloud, or other regulated environment.
  • Experience using logs, metrics, dashboards, and alerts to troubleshoot production services with tools such as Prometheus, Grafana, CloudWatch, CloudTrail, Elastic Stack, or Splunk.
  • Experience supporting CI/CD pipelines and automated production deployments using tools such as GitLab or Jenkins.
  • Experience with highly available, multi-region distributed systems, capacity planning, and disaster recovery.
  • Strong written and verbal communication skills with experience creating runbooks, troubleshooting guides, operational procedures, and audit documentation.
  • Experience participating in an on-call rotation, production incidents, root-cause analysis, or post-incident reviews, with an understanding of Incident Commander responsibilities.

Responsibilities

  • Deploy, operate, and maintain resilient AWS and Kubernetes-based infrastructure and microservices that support the Webex for Government environment.
  • Monitor service health, troubleshoot complex incidents, and help restore service quickly in a 24x7 production environment.
  • Automate repeatable operational work and infrastructure provisioning to improve safety, efficiency, and consistency across deployments.
  • Strengthen platform security and compliance by supporting identity and access management, logging, hardening, vulnerability remediation, and FedRAMP continuous monitoring activities.
  • Collaborate across engineering, security, and compliance teams to identify operational risks, improve reliability, and build documentation, runbooks, and incident reports that make the environment easier to operate.

Benefits

  • medical, dental and vision insurance
  • a 401(k) plan with a Cisco matching contribution
  • paid parental leave
  • short and long-term disability coverage
  • basic life insurance
  • 10 paid holidays per full calendar year, plus 1 floating holiday for non-exempt employees
  • 1 paid day off for employee’s birthday
  • paid year-end holiday shutdown
  • 4 paid days off for personal wellness
  • 16 days of paid vacation time per full calendar year, accrued at rate of 4.92 hours per pay period for full-time employees (non-exempt)
  • flexible vacation time off program (exempt)
  • 80 hours of sick time off provided on hire date and each January 1st thereafter
  • up to 80 hours of unused sick time carried forward from one calendar year to the next
  • Additional paid time away may be requested to deal with critical or emergency issues for family members
  • Optional 10 paid days per full calendar year to volunteer
  • annual bonuses
  • performance-based incentive pay
  • Cisco restricted stock units
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service