Member of Technical Staff, RL Research & Environments

Magic - San Francisco, CA

Hiring: Member of Technical Staff, RL Research & Environments Company: Magic Location: San Francisco, CA Job Posted Time: 2026-09-12 11:07:26 About the job Required Skills: •As a Research Engineer on the RL Research & Environments team, you will design and operate the data, evaluation, and environment systems that improve model capabilities after pre-training. •This role focuses on post-training: identifying capability gaps, building targeted datasets, designing reward signals, and running iterative training loops that measurably improve user-facing behavior. You will own the infrastructure and experimental workflows that connect product priorities to concrete capability gains. •Magic’s long-context models introduce distinct post-training challenges: long-horizon reasoning, sustained coherence over extended trajectories, context-use quality, and tool-augmented behavior. You will build systems that expose failure modes, generate high-signal training data, and enable rapid RL iteration at scale. •This role can evolve into ownership of major capability areas, deeper RL systems work, or broader influence over post-training strategy as Magic scales long-context model performance and reliability. •Design and build post-training datasets using synthetic generation, targeted data collection, and self-play •Implement filtering, scoring, and mixture strategies for RL and post-training corpora •Build and maintain evaluation frameworks that surface long-context failure modes •Design reward signals and training environments for targeted capability improvements Compensation: •$275,000 - $550,000 / year Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!