Researcher, Safety Training, National Security

OpenAI - San Francisco, CA

Hiring: Researcher, Safety Training, National Security Company: OpenAI Location: San Francisco, CA Job Posted Time: 2026-09-12 18:00:16 Target Skills & Keywords : AI, Deep Learning, HBase, Machine Learning, Make, OpenAPI, Reinforcement Learning, SAFe About the job Experience: •4+ years of relevant AI safety research experience, including RLHF, adversarial training, or robustness. Required Skills: •We’re seeking a researcher to train and evaluate models for U.S. government use, with a focus on national security applications. You’ll advance safety post-training and robustness, helping models follow nuanced policies while preserving their usefulness and capabilities. •In This Role, You Will •Research and implement methods for safety training, reinforcement learning, and adversarial robustness. •Develop evaluations, identify model failure modes, and use findings to improve training. •Work with research, engineering, security, and policy partners to support safe, reliable deployment. •You Might Thrive In This Role If You •Bring 4+ years of relevant AI safety research experience, including RLHF, adversarial training, or robustness. •Have a degree in computer science, machine learning, or a related field, and strong deep learning research or engineering skills. •Have experience improving model safety for deployment and enjoy collaborative research. •Are motivated by OpenAI’s mission and the responsible use of AI in safety-critical settings. •Security Requirements •Active TS/SCI clearance or equivalent. •About OpenAI •OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. Compensation: •$380,000 •$500,000 Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!