Software Engineer

NetApp, Inc.Durham, NC
$170,000 - $253,000

About The Position

The Software Engineer will lead a dynamic team responsible for ensuring our critical systems' reliability, performance, and efficiency. This role involves a strategic blend of engineering and operations and requires a strong background in software development, systems engineering, and leadership. This is a pivotal role in our operations, demanding a dedicated individual who excels in a fast-paced and collaborative environment. We invite you to apply if you are driven by system reliability and ready to lead a high-performing team.

Requirements

  • 10 years of experience in Software Development, Platform Engineering, DevOps, or similar roles, with at least 5 years in a lead and/or architect position.
  • This role will be a mix of all 3: Backend Engineering, Site Reliability Engineering and DevOps practices.
  • Experience mentoring geographically dispersed teams.
  • Recommend the appropriate technological approach, team structures, and skill sets.
  • Proficiency in programming languages such as Python, Go, or Java.
  • Extensive experience with cloud services (AWS, GCP, Azure) and container orchestration tools (Kubernetes, Docker).
  • Experience designing and implementing CI/CD pipelines and Configuration Management (Jenkins, Ansible, Terraform).
  • Deliver architectural initiatives that drive and improve efficiency in line with business strategy.
  • Familiarity with distributed systems design patterns using tools such as Kubernetes.
  • Exceptional knowledge of observability tools and setting up architecture for proactive monitoring of the product.
  • Experience in setting up SLOs & SLIs.
  • Proven track record of designing and implementing scalable, high-availability systems.

Responsibilities

  • Lead and mentor a team of Engineers, fostering a culture of continuous improvement and innovation.
  • Collaborate with product and engineering teams to design and implement scalable solutions.
  • Develop and maintain a reliable monitoring and alerting system to detect and mitigate issues proactively.
  • Handle incidents to reduce TTM and TTR consistently.
  • Participate and lead post-mortem analyses to prevent future outages.
  • Manage priorities, projects, and the overall workflow of the SRE team.
  • Ensure compliance with security best practices and company policies.
  • Stay ahead of industry trends and emerging technologies to improve system reliability and performance continuously.
  • Exceptional problem-solving skills and the ability to work under pressure.
  • Excellent communication and team-building skills.

Benefits

  • Health Insurance
  • Life Insurance
  • Retirement or Pension Plans
  • Paid Time Off
  • various Leave options
  • Performance-Based Incentives
  • employee stock purchase plan
  • restricted stocks (RSU’s)
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service