Anthropic is forming a new team focused on steering the world through the next generation of cybersecurity, researching the impacts of advanced AI models on security, publishing this research, and building and deploying necessary defenses. This team is a high-priority initiative within the Frontier Red Team, which studies catastrophic risks and advocates for defenses. The role involves researching offensive and defensive capabilities of AI models like Claude, prototyping defenses, informing AI training and safeguarding to favor defenders, and engaging with the public and government to ensure defenders have a permanent advantage. The Lead will develop a vision for a secure future, communicate it to various stakeholders, and make high-level decisions regarding AI capability distribution, model safeguarding, threat focus, system hardening, and research investments. The role requires building and leading a world-class research team, producing frontier research on quick timelines, maximizing research impact, and deploying defenses.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior