Researcher, Safety Training, National Security

OpenAIWashington, DC
$380,000 - $500,000

About The Position

We’re seeking a researcher to train and evaluate models for U.S. government use, with a focus on national security applications. You’ll advance safety post-training and robustness, helping models follow nuanced policies while preserving their usefulness and capabilities.

Requirements

  • 4+ years of relevant AI safety research experience, including RLHF, adversarial training, or robustness.
  • A degree in computer science, machine learning, or a related field, and strong deep learning research or engineering skills.
  • Experience improving model safety for deployment and enjoy collaborative research.
  • Motivated by OpenAI’s mission and the responsible use of AI in safety-critical settings.
  • Active TS/SCI clearance or equivalent.

Responsibilities

  • Research and implement methods for safety training, reinforcement learning, and adversarial robustness.
  • Develop evaluations, identify model failure modes, and use findings to improve training.
  • Work with research, engineering, security, and policy partners to support safe, reliable deployment.
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service