The Interpretability team at Anthropic works to understand what's actually happening inside trained models and applies techniques to keep frontier AI safe as it rapidly improves. This role is an early hire on a new infrastructure effort within Interpretability, focused on building secure, private, and low-friction access to frontier models for researchers. The work spans four areas: Security (designing secure-by-default environments), Privacy (building data-access patterns for policy adherence), Data & Compute Management (managing research data at petabyte scale and efficient use of accelerator fleets), and Developer Experience (agentic engineering, tooling, and observability). The engineer will be deeply embedded with researchers, understanding their workflows and bridging communication with platform and security teams.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level
Education Level
Associate degree