Anthropic's Safeguards organization builds the policies, evaluations, and enforcement systems that define and hold the limits on how Claude can be used. In this role, you'll own our conventional weapons work. Defining a crisp boundary between acceptable and harmful requests in this domain is difficult, because the underlying components are dual-use: the same capabilities that serve civilian engineering and research can also contribute to a weapons system. Building the threat models, evaluations, and detection systems that hold that boundary is the core of this role. Conventional weapons are increasingly defined by software, and the risks reach well beyond firearms: from a model operating a weapons system or writing guidance code, to the weaponization of dual-use platforms. The role spans every weapon class, up to autonomous systems that select and engage targets without human authorization. You will help define the line between prohibited weapons development and legitimate research and engineering work, and how to operationalize the distinction. This will include translating technical judgment into principles that engineers can implement and enforcement teams can act on.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior