Research Scientist
The Center for AI Safety (CAIS) is a leading research and advocacy organization focused on mitigating societal-scale risks from AI. We address AI’s toughest challenges through technical research, field-building initiatives, and policy engagement, along with our sister organization, Center for AI Safety Action Fund.
As a Research Scientist here, you will lead and execute high-impact research that advances the safety and reliability of frontier AI systems. You'll design and run experiments on large language models, build the tooling needed to train and evaluate models at scale, and turn results into publishable research. You'll collaborate closely with CAIS researchers and external academic and commercial partners, using our compute cluster to run large-scale training and evaluation. The work spans areas like AI honesty, robustness, transparency, and trojan/backdoor behaviors, aimed at reducing real-world risks from advanced AI systems.