PB✓
PBridge

Full-time jobsthe United States

Researcher, Safety Training, National Security

openai · San Francisco · Full-time

About this role

ABOUT THE TEAM

The Safety Training research team aims to fundamentally advance our capabilities for precisely implementing safe behavior in AI models, and to leverage these advances to make OpenAI’s deployed models safe and beneficial. This requires a breadth of new ML research to address the growing set of safety challenges as AI becomes more powerful and used in more settings. Key focus areas include how to train nuanced safety behaviors, how to make the model robust to bad actors, how to address privacy and security risks, and how to make the model trustworthy in safety-critical situations.

We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely.

ABOUT THE ROLE

We’re seeking a researcher to train and evaluate models for U.S. government use, with a focus on national security applications. You’ll advance safety post-training and robustness, helping models follow nuanced policies while preserving their usefulness and capabilities.

IN THIS ROLE, YOU WILL:

- Research and implement methods for safety training, reinforcement learning, and adversarial robustness.

- Develop evaluations, identify model failure modes, and use findings to improve training.

- Work with research, engineering, security, and policy partners to support safe, reliable deployment.

YOU MIGHT THRIVE IN THIS ROLE IF YOU:

- Bring 4+ years of relevant AI safety research experience, including RLHF, adversarial training, or robustness.

- Have a degree in computer science, machine learning, or a related field, and strong deep learning research or engineering skills.

- Have experience improving model safety for deployment and enjoy collaborative research.

- Are motivated by OpenAI’s mission and the responsible use of AI in safety-critical settings.

SECURITY REQUIREMENTS

- Active TS/SCI clearance or equivalent.

About OpenAI

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. 

We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic.

For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement https://cdn.openai.com/policies/eeo-policy-statement.pdf.

Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with

Tired of applying one by one?

Our Career Success Team finds roles in the United States that fit you, tailors your CV to each, and submits the applications — tracked end to end. You just show up to interviews.

We apply, you interview →