AI Red Teamer, CBRNE (Remote)

Handshake - United States

Hiring: AI Red Teamer, CBRNE (Remote) Company: Handshake Location: United States Job Posted Time: 2026-09-09 17:56:26 Employment Type: Full-time Target Skills & Keywords: AI Safety, Red Teaming, CBRNE, WMD Analysis, Threat Assessment, Adversarial Evaluation, LLMs, Jailbreak Techniques, Dual-Use Research, Harm Taxonomies, National Security, Trust & Safety, Python, LLM APIs Experience: - Graduate-level education or equivalent professional experience in a relevant CBRNE field (chemistry, biochemistry, microbiology, virology, nuclear physics, radiochemistry, materials science, munitions/ordnance, chemical engineering, or closely related disciplines) - Strong hands-on experience using multiple LLMs (ChatGPT, Claude, Gemini, open-source models, etc.) - Nice to have: Experience in threat assessment, WMD analysis, intelligence analysis, or arms control verification - Nice to have: Experience in red teaming, penetration testing, or structured adversarial evaluation in any context - Nice to have: Prior work in trust and safety, content moderation, or AI evaluation Required Skills: - Design technically grounded adversarial prompts that test whether models provide meaningful uplift toward CBRNE threats - Evaluate model outputs for technical accuracy, assessing whether responses contain genuinely dangerous information versus superficial or publicly available knowledge - Probe dual-use knowledge boundaries, testing how models handle queries that blend legitimate scientific, medical, or industrial use cases with potential weapons applications - Test multi-step and multi-turn attack chains that simulate how a motivated actor might extract dangerous information incrementally - Score model responses against structured harm taxonomies and severity rubrics calibrated to real-world risk - Document findings with clear technical reasoning, including what a response gets right, what it gets wrong, and why the failure matters - Identify and articulate the difference between information that is freely available in open literature and information that constitutes genuine uplift beyond baseline - Contribute to the development and refinement of CBRNE-specific evaluation frameworks and threat models - Collaborate with other red teamers, AI researchers, and policy teams to translate findings into actionable model improvements - Stay current on evolving model capabilities, jailbreak techniques, and relevant developments in your domain - Ability to evaluate the technical accuracy and real-world consequence of model outputs in your domain - Understanding of dual-use research concerns and the distinction between open-source knowledge and operationally significant uplift - Creative, adversarial problem-solving skills - Clear and precise written communication, including the ability to explain technical risk to non-specialist audiences - Strong ethical judgment and the ability to separate adversarial thinking from personal values - Self-directed, collaborative, and comfortable in feedback-heavy environments Qualifications: - Graduate-level education or equivalent professional experience in a relevant CBRNE field (chemistry, biochemistry, microbiology, virology, nuclear physics, radiochemistry, materials science, munitions/ordnance, chemical engineering, or closely related disciplines) - Deep subject matter expertise in at least one CBRNE domain - Strong ethical judgment and the ability to think like a sophisticated threat actor while operating within a structured evaluation framework - Nice to have: Active or prior security clearance (Secret, Top Secret, or SCI) - Nice to have: Background in biosafety/biosecurity, chemical safety, nuclear nonproliferation, or explosive ordnance disposal - Nice to have: Familiarity with relevant regulatory frameworks (CWC, BWC, IAEA safeguards, ATF regulations, Export Administration Regulations) - Nice to have: Familiarity with Python or scripting languages, LLM APIs, or evaluation tooling - Nice to have: Published research or professional presentations in a relevant CBRNE domain Compensation: - Base pay range: $65.00/hr - $158.00/hr Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don’t miss this opportunity to join a forward-thinking team!