The Cloud & AI organization accelerates Microsoft's mission to secure digital technology platforms, devices, clouds, and AI systems across customers' heterogeneous environments and Microsoft's internal estate. Our culture is centered on a growth mindset, inspiring excellence, and helping teams and leaders bring their best every day. Microsoft Red Team (MRT) emulates real-world advanced persistent threats against Microsoft, external customers, and frontier AI systems. As AI systems become increasingly capable at reasoning, coding, tool use, exploitation, and autonomous execution, understanding when and how those systems materially increase offensive cyber capability is an important component of Microsoft's AI security mission. MRT Strategic AI is seeking an AI Security Engineer focused on Cyber Capability Evaluation. This is a hands-on technical role for implementing and executing evaluations of frontier AI models and agentic systems in realistic, controlled cyber environments. The engineer will work at the intersection of offensive security, AI agents, model evaluation, and experimental engineering to build and maintain evaluation tasks, run experiments, analyze model behavior, and determine whether observed results represent meaningful cyber capability, an evaluation artifact, or a limitation in the test environment. The role requires the ability to work independently within established technical direction, solve engineering and security problems, and contribute improvements to evaluation methods, cyber ranges, scoring, telemetry, and containment. The right candidate combines offensive-security experience with experimental discipline, programming skills, and an interest in emerging AI capabilities.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Entry Level