Data Scientist
Peraton - Red Bank, NJ
Hiring: Data Scientist Company: Peraton Location: Red Bank, NJ Job Posted Time: 2026-09-16 18:06:20 Employment Type: Contract Target Skills & Keywords : AI, Airflow, Dagster, Data Pipeline, Deep Learning, Machine Learning, Neo4j, NumPy, Pandas, Prefect, PyTorch, Python, SQL, SciPy, Security Clearance, scikit-learn About the job Experience: •5+ years of applied data science or data engineering experience (or MS with 3+ years) with a record of delivering data pipelines and analyses that other engineers and researchers depend on Required Skills: •The ideal candidate is a strong applied data scientist with an interest in symbolic and structured representations of knowledge. Experience with formal methods and domain-specific languages (DSLs) is desired but not required. •Design, build, and populate the research team's effort knowledge bases, and maintain it under version control with provenance tracking •Develop knowledge extraction pipelines (rule-based, statistical, and ML-assisted) that convert unstructured and semi-structured sources into structured, queryable knowledge; establish quality metrics and validation procedures for extracted content •Engineer the metadata, annotation schema, and packaging for program datasets (synthetic, simulated, and real collections) so that datasets are reproducible, well documented, and deliverable on the program data sharing schedule •Perform exploratory and inferential analysis on training data, simulation output, and evaluation results; build dashboards and reports that show where generated waveforms succeed or fail against objectives, and feed findings back to AI model researchers and engineers •Contribute data and analysis sections to design reviews, monthly status reports, and dataset documentation; coordinate with academic subcontractors on shared data and knowledge resources Qualifications: •Bachelor's Degree or higher in Computer Science, Statistics, Mathematics, Electrical Engineering, or a related technical field •Strong Python proficiency including the scientific stack (NumPy, pandas, SciPy, scikit-learn) and experience with at least one deep learning framework (PyTorch preferred) •Applied hands-on capability in knowledge graphs, ontologies, or structured knowledge representation (RDF/OWL, property graphs such as Neo4j, or equivalent) and with querying and validating them (SPARQL, Cypher, SHACL, or similar) •Solid statistical foundations: experimental design, hypothesis testing, uncertainty quantification, and the ability to explain results to technical and non-technical audiences •Demonstrated capacity to produce clear documentation including data dictionaries, dataset cards, and analysis reports •Background in symbolic AI or neuro-symbolic methods: logic programming, constraint solving, rule engines, or program synthesis •Operational familiarity with digital signal processing and communications fundamentals (modulation, filtering, coding, channel effects) or with GNU Radio and software-defined radio data formats (I/Q sample handling, SigMF or similar metadata standards) •Prior work on IARPA, DARPA, or similar government research programs, including data sharing plans, privacy protection plans, and delivery of datasets to independent T&E teams •MS or PhD in Computer Science, Statistics, Electrical Engineering, or a related technical field Compensation: •$146,000 - $234,000 / year •Depending on the position, employees may be eligible for overtime, shift differential, and a discretionary bonus in addition to base pay Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!