Member of Technical Staff - Multi-Modal, Vision

Liquid AI - San Francisco, CA

Hiring: Member of Technical Staff - Multi-Modal, Vision Company: Liquid AI Location: San Francisco, CA Job Posted Time: 2026-09-12 12:46:35 Target Skills & Keywords : Computer Vision, Data Pipeline, Deep Learning, GitHub, Hugging Face, Python, Reinforcement Learning About the job Required Skills: •The VLM team builds vision-language models that run on-device, under tight latency and memory constraints, without sacrificing quality. We have released four best-in-class models and we're just getting started. Qualifications: •Applied hands-on capability in training or evaluating VLMs with demonstrated experimental rigor. •Demonstrated capacity to turn research ideas into scalable implementations, refine and iterate through hypotheses. •Proficiency in Python and at least one deep learning framework. •M.S. or Ph.D. in Computer Science, Mathematics, or a related field; or equivalent industry experience. •Building or optimizing multimodal training or data pipelines. •Multimodal post-training experience (SFT, preference optimization, RL-style methods). •Dataset design and data quality expertise (quality and diversity assessment, long-tail mining). •Prior open-source contributions (code, data, models) on GitHub or Hugging Face. •Published research at top AI conferences. •What Working Here Might Look Like Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!