Software Engineer, Infrastructure, Interpretability

AnthropicSan Francisco, CA
Hybrid

About The Position

The Interpretability team at Anthropic works to understand what's actually happening inside trained models and applies techniques to keep frontier AI safe as it rapidly improves. This role is an early hire on a new infrastructure effort within Interpretability, helping to define its charter. The job is to build the paved path that makes deep model access secure by default, private by design, and low-friction for every researcher. The work spans four areas: Security (design secure-by-default environments and access patterns), Privacy (build data-access patterns that ensure policy adherence), Data & Compute Management (manage research data at petabyte scale and make efficient use of large accelerator fleets), and Developer experience (agentic engineering, tooling and observability that keep researchers moving fast). In this role, you’ll be deeply embedded alongside Interpretability Researchers to understand their workflows, building your understanding of the research as you go. You’ll also bridge communication with Anthropic’s wider platform and security teams. Every hour of researcher friction you remove is multiplied across the whole organization, and the infrastructure you build sets the pace at which interpretability results reach real safety decisions.

Requirements

  • Highly proficient in at least one programming language (e.g., Python, Rust, Go, Java) and productive with Python
  • Significant experience building and operating secure and scalable software infrastructure - cloud systems, distributed systems, or developer tooling
  • Strong cross-functional communication skills - equally at home working with researchers and with platform and security teams
  • Extremely curious about unfamiliar domains
  • Strong ability to prioritize the most impactful work and are comfortable operating with ambiguity and questioning assumptions
  • Curious about interpretability research and its role in AI safety (though no research experience is required!)
  • Care about the societal impacts and ethics of your work

Nice To Haves

  • Experience with cloud infrastructure (e.g. GCP or AWS), Kubernetes, networking and infrastructure-as-code
  • Security engineering experience: identity / auth / access management, sandboxing, red teaming
  • Experience with data warehousing, large-scale storage systems, and data lifecycle management - especially for research
  • Experience with compute schedulers and accelerator fleet management
  • Experience building developer productivity tooling and observability stacks
  • Experience building tooling to accelerate research teams

Responsibilities

  • Design, build, and own shared infrastructure for Interpretability - research environments, data systems, and compute tooling that researchers rely on daily
  • Lead cross-team efforts with our agentic engineering, security, compute, and storage platform teams, so that company-wide solutions serve research needs
  • Discover and resolve major organization-wide developer experience issues
  • Help take interpretability methods from research code to dependable audit pipelines

Benefits

  • competitive compensation and benefits
  • optional equity donation matching
  • generous vacation and parental leave
  • flexible working hours
  • a lovely office space in which to collaborate with colleagues
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service