Site Reliability Engineer- REMOTE - Onsite Training

NTT DATA ServicesMemphis, TN
$87,120 - $151,250Remote

About The Position

NTT DATA strives to hire exceptional, innovative and passionate individuals who want to grow with us. If you want to be part of an inclusive, adaptable, and forward-thinking organization, apply now. We are currently seeking a Site Reliability Engineer- REMOTE - Onsite Training to join our team in Memphis, Tennessee (US-TN), United States (US).

Requirements

  • 8+ years of experience in DevOps/SRE environments, including substantial experience with CI/CD platforms, containerization technologies, and Kubernetes-based deployments.
  • 5+ years’ experience in troubleshooting using APM (Application Performance Management) tools like Datadog and Dynatrace.
  • 4+ years' experience with Unix/Linux shell scripting and supporting NoSQL databases such as Couchbase.
  • 3+ years' experience with static code analysis tools such as Checkmarx and SonarQube.
  • Experience with software and platform design, implementation, CI/CD pipelines, EKS, Bamboo, Jenkins, Docker, Splunk, and Datadog or similar monitoring tools.

Nice To Haves

  • Strong log analysis using tools like Splunk.
  • Experience in working with Nexus Repository.
  • Ability to diagnose, troubleshoot complex issues.
  • Demonstrated advanced ability in one or more of the required technical areas.
  • Demonstrated commitment to continuous improvement and operational excellence.
  • Experience with ServiceNow and Jira for change management.
  • OWASP knowledge
  • Project lead experience

Responsibilities

  • Implement and support CI/CD tools and pipelines across the organization (GitLab preferred).
  • Monitor production/non-production systems and help solve problems, using tools including Datadog.
  • Maintain containerized applications with Docker or similar technologies.
  • Work with the product and development teams to continue to scale our application and infrastructure while ensuring performance and high availability.
  • Collaborate with the development team on defining Service Level Indicators (SLIs) that represent the health of their service.
  • Develop systems and software that increase site reliability and performance and be part of building SRE competency within the Digital organization.
  • Help evolve the CI/CD pipeline leveraging SLIs and other monitorability works. Help with automating testing strategies to determine quality gates for production deployment.
  • Build long term automation solutions using scripting and programming (Groovy, Shell, Python, Terraform, or Java/JavaScript).
  • Provide technical leadership and implementation guidance for UI, APIs, and microservices.
  • Participate in a scheduled on-call rotation supporting production systems.

Benefits

  • medical, dental, and vision insurance with an employer contribution
  • flexible spending or health savings account
  • life and AD&D insurance
  • short and long term disability coverage
  • paid time off
  • employee assistance
  • participation in a 401k program with company match
  • additional voluntary or legally-required benefits
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service