Data Scientist I - Python and SQL
The Data Scientist I at Confidential Client supports diverse research projects by applying statistical, machine learning, and AI methods to complex datasets including survey and administrative data. This hybrid role involves developing reproducible analytical workflows, building data pipelines, and creating data products to enhance research and operational decision-making. The position emphasizes collaboration with multidisciplinary teams and the application of emerging technologies such as generative AI and large language models.
About Careertakes
👉 Important disclosure: Careertakes is a third-party recruiting platform supporting this hiring process. If selected, you will be employed directly by our client, Scientific & QA.
Applicants for this role may also receive access to additional matched opportunities through the Careertakes platform.
Overview
Confidential Client is seeking a Data Scientist I to join a Statistics & Data Science team supporting survey research, evaluation studies, and broader social‑science and policy research. This early‑career role combines data engineering, reproducible analytic workflows, statistical modeling, and applied machine learning (including NLP and LLM evaluation) to turn complex structured and unstructured data into reliable, actionable insights.
Responsibilities
- Apply statistical, computational, and machine‑learning methods to support survey research, evaluation, and other social science projects.
- Build and maintain reproducible analytic workflows: version control (Git), environment management, automated testing, and peer code review.
- Develop data pipelines that integrate, transform, and analyze structured and unstructured data (surveys, administrative records, commercial files, text).
- Create data products, dashboards, and analytical applications to support decision‑making.
- Design, evaluate, and help deploy ML/AI solutions, including natural language processing and LLM experimentation, in collaboration with multidisciplinary teams.
- Write production‑quality code (Python, and R where needed) to extract, link, and analyze datasets.
- Work with relational databases, cloud platforms, and large‑scale compute environments.
- Follow data disclosure limitation and privacy best practices for public and restricted data releases.
- Perform other duties as assigned.
Required Qualifications
- Bachelor’s degree in computational social science, data science, statistics, computer science, or a related quantitative field.
- Approximately 4 years of relevant experience (including graduate research, internships, or applied professional work).
- Demonstrated proficiency producing production‑quality Python code.
- Strong SQL and relational database experience.
- Experience building software, data pipelines, or analytical applications.
- Hands‑on experience with statistical analysis and machine learning on real‑world datasets.
- Familiarity with supervised and unsupervised ML methods.
- Experience with Git and collaborative development workflows.
- Strong problem‑solving, communication, and technical writing skills; able to explain technical concepts to non‑technical audiences.
Preferred Qualifications
- Experience with large‑scale data platforms (Databricks, Spark, Hadoop, Hive).
- Experience developing, evaluating, or deploying ML/AI solutions, including work with large language models and cloud environments (AWS, Azure, GCP).
- Familiarity with data disclosure limitation (SDL), data privacy methods, and risk assessment for public releases.
- Experience with R or SAS.
- Prior work analyzing healthcare, survey, administrative, or other large observational datasets (e.g., claims data).
- Experience with record linkage and data integration across sources.
Salary & Benefits
- Salary range: $90,000 - $100,000 per year.
- Regular staff are eligible for a comprehensive benefits program, including subsidized health insurance (effective day one), dental and vision, defined contribution retirement and optional 403(b), group life and disability insurance, generous paid time off, paid parental leave, tuition assistance, and an Employee Assistance Program (EAP).
- Confidential Client publishes salary ranges and considers experience, competencies, and qualifications when placing candidates within the posted range.
Work Location & Eligibility
- Authoritative location: Chicago, IL. This is a hybrid role based in the Chicago office with an expectation of at least six days per month in the office.
- Candidates must be eligible to work in the U.S.; Confidential Client is unable to offer visa sponsorship for this position.
Equal Opportunity & Hiring Transparency
Careertakes and our client are Equal Opportunity Employers committed to building a diverse and inclusive workforce. We prohibit discrimination or harassment of any kind. To support a fair and efficient hiring process, AI tools may be used to assist with application review or resume screening. These tools do not replace human decision-making. Final hiring decisions are made by people.
If you have questions about how your data is used, please contact us directly.
Negotiate a higher salary! Check the salary ranges for this job type in your area.
View My Salary Range