Software Engineer, Site Reliability

Hebbia AINew York, NY
Onsite

About The Position

Hebbia is an AI platform for investors and bankers that generates alpha and drives upside. Founded in 2020, Hebbia powers investment decisions for major financial institutions. Their flagship product, Matrix, delivers industry-leading accuracy, speed, and transparency in AI-driven analysis, managing over $30 trillion in assets globally. Hebbia provides intelligence that gives finance professionals a competitive edge by uncovering hidden opportunities and accelerating decisions. The company transforms how capital is deployed, risk is managed, and value is created across markets, offering a competitive advantage that drives performance and market leadership. The role is for a Site Reliability Engineer who thinks like a software engineer first. This individual will own critical production systems end-to-end, focusing on designing, building, and improving them. Responsibilities include writing production-quality code for reliability at scale, embedding with product engineering teams to influence architecture, and building essential internal tooling. This is not a ticket-driven operations role; the focus is on writing code for service instrumentation, performance bottleneck elimination, deployment platform development, and translating incident learnings into architectural improvements.

Requirements

  • 5+ years software development with a track record of writing, shipping, and maintaining production services, not just operating infrastructure
  • Production-grade proficiency in at least one systems or backend language: Go, Python, C++, or Rust
  • Proven experience as a Production Engineer, SRE, or software engineer with a deep infrastructure focus, comfortable owning services end-to-end across the full stack
  • Deep understanding of distributed systems
  • Container orchestration expertise and hands-on experience debugging complex distributed failures in production
  • Working knowledge of OS-level concepts
  • Cloud platform fluency (AWS preferred)
  • Experience in building and maintaining observability stacks
  • Strong CI/CD pipeline expertise and a track record of improving developer velocity without sacrificing safety

Nice To Haves

  • Background at a company with a Production Engineering or software-focused SRE culture is a strong plus
  • Experience building platforms for AI/ML workloads or high-throughput document processing pipelines is a plus

Responsibilities

  • Own critical production services end-to-end, from design and code review through deployment, operation, and incident response
  • Profile, benchmark, and rewrite hot paths to eliminate bottlenecks as Hebbia scales
  • Lead incident response and drive post-mortem culture, translating findings into code changes and architectural improvements rather than runbooks
  • Design and build observability frameworks from scratch, writing custom instrumentation, alerting logic, and debugging tooling that surfaces production issues before customers feel them
  • Define and enforce SLOs across platform services and build the feedback loops that keep engineering teams accountable to them
  • Own capacity planning and cost efficiency: model growth, right-size infrastructure, and write automation that prevents over-provisioning and resource exhaustion
  • Build robust, well-tested internal platforms and deployment tooling held to the same engineering standards as customer-facing code
  • Own and continuously improve CI/CD systems so engineering teams can ship safely and quickly
  • Embed with product engineering teams as a peer software engineer, contributing directly to production codebases and co-designing systems for reliability from the start
  • Partner on infrastructure security through threat modeling, hardening, and automated compliance tooling

Benefits

  • Unlimited PTO
  • Medical + Dental + Vision + 401K
  • Catered lunch daily + Doordash dinner credit if you ever need to stay late
  • 3 months non-birthing parent parental leave
  • 4 months birthing parent parental leave
  • $15k lifetime fertility benefits
  • Competitive equity package with unmatched upside potential
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service