Postdoctoral Research Associate, Machine Unlearning and Model Editing for AI Biosecurity

UVA HealthCharlottesville, VA
$60,000 - $75,000Onsite

About The Position

This position develops and evaluates machine unlearning and model editing methods that selectively reduce hazardous biological capabilities in AI systems while preserving beneficial scientific functions. The researcher reports to Assistant Professor Tom Hartvigsen and will work closely with other SDS faculty members including Chirag Agarwal, and Stephen Turner, and works with faculty in interpretability and with a laboratory partner that leads adversarial red teaming. The role centers on implementing, innovating, and comparing model editing and unlearning methods, measuring safety--utility tradeoffs against both benchmarks and realistic task batteries, and leading technical development of an open evaluation suite for AI biosecurity. Strong familiarity with biology and biosecurity is important, as the work targets biological capabilities and connects to a human-subjects evaluation running in parallel.

Requirements

  • Doctoral degree (PhD or equivalent) in data science, computer science, machine learning, or a related field, completed at the time of hire
  • Strong programming in Python and hands-on experience with modern ML frameworks such as PyTorch and Hugging Face Transformers
  • Track record of publications in machine learning, natural language processing, and/or biosecurity
  • Demonstrated experience training, finetuning, or post-training for large language models
  • Software engineering practices that support reproducible and reusable research tools

Nice To Haves

  • Understanding of biology, biosecurity, or dual-use research considerations
  • Experience with machine unlearning, model editing, or related capability-mitigation methods
  • Experience with mechanistic interpretability or representation analysis
  • Familiarity with adversarial robustness, red-teaming, or jailbreak evaluation
  • Experience releasing and maintaining open-source ML evaluation tooling
  • Familiarity with secure computing environments and controlled-access model arrangements

Responsibilities

  • Implement and compare machine unlearning and model editing methods, including gradient-based fine-tuning, representation-level edits, and inference-time steering
  • Design and run experiments that measure how interventions affect benchmark scores and real-world task performance, producing safety-utility curves
  • Develop adversarial testing protocols with the laboratory partner, including prompt-based jailbreaks, fine-tuning recovery, and ensemble attacks
  • Lead engineering of the open-source UBS-Bio evaluation suite, including baselines, metrics, and documentation
  • Support interpretability analyses that identify which model representations encode hazardous versus beneficial capabilities
  • Prepare and present manuscripts and publish and maintain reproducible code releases

Benefits

  • UVA Health Plan: the choice between 3 different health plans
  • Vision Coverage
  • Dental Plan
  • Benefit Savings Plans
  • Life Insurance
  • Disability Benefits
  • Paid Time Off: starting with 22 days of time off per year, 12 or more holidays, 8 weeks parental leave
  • Use of up to $5250 per calendar year towards a for-credit degree program or for-credit certificate program
  • Use of up to $2000 of the total $5250 noted above per calendar year for professional development including job-related training, conferences, and initial certificate exams.

Stand Out From the Crowd

Upload your resume and get instant feedback on how well it matches this job.

Upload and Match Resume

What This Job Offers

Job Type

Full-time

Career Level

Entry Level

Education Level

Ph.D. or professional degree

© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service