Staff Production Engineer, Database Infrastructure

GitHub, Inc.UNAVAILABLE, UNAVAILABLE
Remote

About The Position

GitHub is looking for a Staff Production Engineer to help scale our data platform to millions of developers. We are software engineers who specialize in reliability, working with technical partners, contributing to design reviews, writing SDKs and tooling to build against, and shaping how our data platform is used to prevent scaling problems before they reach production.

Requirements

  • 9+ years experience in Software Engineering, Computer Science, or related technical discipline with proven experience maintaining and delivering production software coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, Go, Ruby, Rust, or Python OR Associate’s Degree in Computer Science, Electrical Engineering, Electronics Engineering, Math, Physics, Computer Engineering, Computer Science, or related field AND 8+ years experience in Software Engineering, Computer Science, or related technical discipline with proven experience maintaining and delivering production software coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, Go, Ruby, Rust, or Python OR Bachelor's Degree in Computer Science or related field AND 7+ years experience in Software Engineering, Computer Science, or related technical discipline with proven experience maintaining and delivering production software coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, Go, Ruby, Rust, or Python. OR Master's Degree in Computer Science, Electrical Engineering, Electronics Engineering, Math, Physics, Computer Engineering, Computer Science, or related field AND 5+ years experience in Software Engineering, Computer Science, or related technical discipline with proven experience maintaining and delivering production software coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, Go, Ruby, Rust, or Python. OR Doctorate in Computer Science, Electrical Engineering, Electronics Engineering, Math, Physics, Computer Engineering, Computer Science, or related field AND 3+ years experience in Software Engineering, Computer Science, or related technical discipline with proven experience maintaining and delivering production software coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, Go, Ruby, Rust, or Python. OR equivalent experience.
  • 3+ years experience operating large-scale distributed systems in production, including participation in an on-call rotation.

Nice To Haves

  • Excitement about building, operating, and maintaining resilient, scalable systems that impact a global community of users with the ability to break down complex systems into manageable components.
  • Experience partnering with product and feature teams to improve the reliability and scalability of what they ship.
  • Ability to influence engineering decisions and proactively engage in system design conversations.
  • Experience running stateful services on managed cloud data stores, specifically Azure Cosmos DB and Azure SQL Database, or equivalents such as DynamoDB, Aurora, or Cloud Spanner, with a focus on partition and index design, consistency tradeoffs, and throughput provisioning.
  • Experience diagnosing and resolving application-level scalability problems: N+1 query patterns, hot partitions, unbounded fan-out, cache stampedes, and data access patterns that don't scale.
  • Experience building internal tooling, SDKs, or libraries used by other engineering teams, written in production-grade Go, Python, Ruby, or Rust.
  • Familiarity with the failure modes of large-scale systems, both in the application and in the platform beneath it e.g. cascading failures, retry storms, thundering herds, partial outages, throttling and quota limits, control plane outages, noisy neighbors, and the patterns that mitigate them.
  • Experience contributing to cloud migrations of live services, including dual-write and backfill strategies, traffic shifting, and rollback planning.
  • Effective communication skills and willingness to pair on problems, brainstorm in public, and enthusiastically engage with your teammates in group problem solving.

Responsibilities

  • Partner with product and feature teams by joining design reviews and shaping how they model, access, and scale their data so the applications they build are performant, available, and operable at GitHub's scale
  • Design and ship SDKs, client libraries, and the tooling applications are built on
  • Define and maintain SLOs and operational standards for GitHub's database platform, and use them to drive prioritization
  • Write technical documentation and advocate for the health and quality of the systems the team builds
  • Participate in an on-call rotation and respond to incidents as needed
  • Contribute to plans for disaster recovery, load shedding, and regional failover

Benefits

  • competitive pay
  • generous learning and growth opportunities
  • excellent benefits
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service