Data Scientist - Identity & KYC
Glocomms - San Francisco County, CA
Hiring: Data Scientist - Identity & KYC Company: Glocomms Location: San Francisco County, CA Job Posted Time: 2026-09-10 14:39:09 Target Skills & Keywords : AWS, Data Lake, Data Pipeline, Data Warehouse, ETL, Linux, Machine Learning, PySpark, PyTorch, Python, R, SQL, Scala, Spark, TensorFlow, XGBoost, scikit-learn About the job Experience: •Master's degree with 2+ years of relevant experience, PhD with 1+ years of experience, or equivalent industry experience in data science, analytics, or machine learning. •2+ years of relevant experience, PhD with 1+ years of experience, or equivalent industry experience in data science, analytics, or machine learning. Required Skills: •A high-growth data and analytics organization is seeking a Data Scientist to join its Big Data R&D team. This team is responsible for developing large-scale identity graph and entity resolution capabilities that support compliance, risk, and fraud-related products. •Contribute to the development of machine learning, statistical, data mining, and graph-based algorithms for identity resolution, anomaly detection, and large-scale data analysis. •Analyze complex datasets to improve identity matching, record linkage, and entity resolution capabilities. •Build and maintain scalable data processing pipelines, including ETL workflows, feature engineering, data normalization, and quality checks. •Support model development through feature engineering, exploratory analysis, error investigation, and experimentation. •Evaluate new internal and third-party data sources by assessing quality, coverage, and impact on analytical models. •Develop and maintain SQL, Python, and related data workflows for extraction, transformation, and validation processes. •Provide analytical support for compliance, risk, and operational teams through investigations, reporting, dashboards, and deep-dive analyses. Qualifications: •Proficiency in Python, Scala, or another data-focused programming language. •Strong SQL skills and experience working with large-scale data warehouse or data lake environments. •Operational familiarity with Linux/Unix environments and cloud platforms, particularly AWS services. •Understanding of supervised and unsupervised machine learning techniques, clustering methods, similarity metrics, and model evaluation approaches. •Demonstrated capacity to frame ambiguous problems, identify key questions, and iterate quickly based on feedback. •Operational familiarity with identity resolution, entity matching, compliance, fraud, risk, or trust and safety use cases. •Exposure to automated data pipelines, experimentation frameworks, and production analytics environments. Compensation: •$140,000 - $190,000 / year Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!