Gray Swan is on a mission to empower the world to use AI safely and securely. We evaluate AI models for the leading frontier labs along with building real-time threat detection and adaptive adversarial red teaming agents for teams deploying AI. We're a team of approximately 50 people, well-funded, growing quickly. Our work directly influences how the world deploys AI agents and systems at scale. Come build and lead Gray Swan's cyber safety capability from the ground up, serving as the technical authority on AI-enabled cyber risk across red-teaming, evaluation, benchmarking, defenses, and safety infrastructure development. You'll help define how frontier AI systems are evaluated for offensive cyber capabilities while partnering with leading AI labs to reduce real-world security risks. This role sits at the intersection of offensive security, AI safety, and machine learning. You'll transform deep cybersecurity expertise into scalable evaluation methodologies, safety infrastructure, and automated defenses that help establish industry standards for frontier model security. If you have deep expertise in offensive cybersecurity, vulnerability research, or adversarial AI security, experience evaluating frontier models, and are driven to reduce catastrophic cyber risks from increasingly capable AI systems, we'd love to hear from you.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed