Data Scientist - Computational Research & ML
Core
Apply statistical, computational, and machine learning methods to survey research, evaluation studies, and social science projects to support data-informed decisions.
Role type
Data Scientist I (Computational Research & ML)
Builds
Reproducible analytical workflows, data pipelines, dashboards, and AI-enabled analytical applications.
Domain
Social science research, survey analysis, and public data privacy
Deliverable
dashboards & analysis
Required skills
Python, SQL, statistical analysis, machine learning, data pipeline development, Git, relational databases
Preferred skills
Large-scale data platforms (Databricks, Spark, Hadoop), cloud environments (AWS, Azure, GCP), LLMs, statistical disclosure limitation, R/SAS, healthcare claims analysis, record linkage
Technologies
Python, R, SQL, AWS, Azure, GCP, Databricks, Spark, Hadoop, Hive, Git, SAS
Responsibilities
Develop data pipelines to integrate and transform structured and unstructured data; Apply machine learning and NLP methods to design and deploy AI-enabled applications; Create data products and automated tools; Protect data privacy using statistical disclosure limitation practices; Collaborate with multidisciplinary teams on research projects.
Seniority
Junior, hands-on IC