Anjali Gopal
Source-listed: Research Scientist, Bioengineering · Anthropic
- Language models
- Evaluation
Selected work
4- Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red TeamingAuthor · 2025
- The WMDP Benchmark: Measuring and Reducing Malicious Use With UnlearningContributor · 2024
- Will releasing the weights of future large language models grant widespread access to pandemic agents?Contributor · 2023
- Best Practices for Germicidal Ultraviolet-C Dose Measurement for N95 Respirator DecontaminationContributor · 2021