Backed by Greylock, Insight and Khosla.
About the role
We’re seeking a researcher to train and evaluate models for U.S. government use, with a focus on national security applications. You’ll advance safety post-training and robustness, helping models follow nuanced policies while preserving their usefulness and capabilities. Research and implement methods for safety training, reinforcement learning, and adversarial robustness.
What they're looking for
- Active TS/SCI clearance or equivalent
More about this role
The Safety Training research team aims to fundamentally advance our capabilities for precisely implementing safe behavior in AI models, and to leverage these advances to make OpenAI’s deployed models safe and beneficial. This requires a breadth of new ML research to address the growing set of safety challenges as AI becomes more powerful and used in more settings. Key focus areas include how to train nuanced safety behaviors, how to make the model robust to bad actors, how to address privacy and security risks, and how to make the model trustworthy in safety-critical situations.
We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely.
We’re seeking a researcher to train and evaluate models for U.S. government use, with a focus on national security applications. You’ll advance safety post-training and robustness, helping models follow nuanced policies while preserving their usefulness and capabilities.
Research and implement methods for safety training, reinforcement learning, and adversarial robustness.
Develop evaluations, identify model failure modes, and use findings to improve training.
Work with...
Browse similar: AI jobs · AI startup jobs · Startup jobs · San Francisco Bay Area