Generative AI Engineer, Senior Staff/Manager (LLM/VLM)

Qualcomm - Ho Chi Minh City Metropolitan Area

Hiring: Generative AI Engineer, Senior Staff/Manager (LLM/VLM) Company: Qualcomm Location: Ho Chi Minh City Metropolitan Area Job Posted Time: 2026-09-16 12:39:02 Target Skills & Keywords : LLM, Machine Learning, PyTorch, Python About the job Experience: •3+ years of hands-on experience designing, training, or optimizing LLMs, VLMs, or foundation models in real-world settings. •4+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience. •3+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience. •2+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience. Required Skills: •Train and fine-tune large-scale LLMs and VLMs for multi-modal and agentic tasks. •Build end-to-end agentic AI prototypes for both on-device and cloud deployment. •Drive model efficiency through quantization, pruning, and distillation — without sacrificing accuracy. •Collaborate closely with hardware, software, and systems teams to bring research to real-world products. •Implement experiments in Python and PyTorch; manage benchmarks across large-scale distributed setups. •Share knowledge and best practices across teams, supporting the growth of junior engineers organically. •Your work will power next-generation AI experiences on •Platforms — from the phone in your pocket to autonomous vehicles and industrial robots. You'll be part of a world-class team that turns research papers into products used by billions. Qualifications: •Bachelor's, Master's, or PhD in Computer Science, Machine Learning, or related field — what matters most is what you've built and shipped. •Strong grasp of generative AI, multi-modal architecture, and training strategies at scale. •Proficiency in Python and PyTorch; experience with distributed training frameworks (e.g., DeepSpeed, FSDP, or equivalent). •Track record applying quantization, pruning, or distillation to reduce model footprint without degrading quality. •PhD in Computer Science or Machine Learning •Operational familiarity with hardware-aware AI optimization and deploying models on edge devices. •Publications at top ML/AI venues — NeurIPS, ICML, CVPR, ACL, etc. •Bachelor's degree in Computer Science, Engineering, Information Systems, or related field and 4+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience. •Master's degree in Computer Science, Engineering, Information Systems, or related field and 3+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience. •PhD in Computer Science, Engineering, Information Systems, or related field and 2+ years of Hardware Engineering, Software Engineering, Systems Engineering, or related work experience. Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!