Data Engineer

Child, Inc.
New York, NY, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$119,000.0 - $150,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Data Analysis Microsoft Azure Cloud Computing Data Infrastructure Extract Transform Load (ETL) Data Transformation Data Security Data Structures Data Visualization Linux
+19 more
Github Python (Programming Language) MATLAB Machine Learning NoSQL SciPy SQL Databases Virtualization Technology Pytorch Model Validation Containerization Scikit Learn Kubernetes Information Technology Data Analytics Terraform Software Version Control Data Pipelines Docker

Job description

As part of the Center for Data Analytics, Innovation, and Rigor team, you will report to Rubric Engineering and Measurement Specialist. You will develop infrastructure to support large-scale AI evaluation frameworks. You will design scalable data pipelines for generating and processing synthetic data, implement secure data storage solutions, and create infrastructure for real-time model evaluation and monitoring. You will use common frameworks, platforms, and languages, such as Python, SQL, GitHub, containerization tools (e.g., Docker, Kubernetes), and cloud computing infrastructures (e.g., AWS, Azure) to build robust and scalable data infrastructure that support our AI research initiatives.

This is an exempt, full-time, hybrid position located in our NYC headquarters office or other relevant location. This position requires a minimum of four (4) days per week in the office, on a schedule determined by your supervisor. The in-office requirement and schedule are subject to change based on the needs of the program and the organization.

You Will:

  • Create and maintain scalable data pipelines for efficient storage and retrieval of multimodal data, with particular emphasis on clinical, natural language, and multi-turn response data.
  • Create pipelines for data transformation, preprocessing, and management. * Ensure data quality, security, and compliance with privacy regulations for handling sensitive data.
  • Perform quality assurance of pipelines/processes to maintain integrity throughout the data lifecycle.
  • Create interactive visualizations and dashboards to communicate data insights and pipeline performance metrics.
  • Write documentation and relevant text for scientific, clinical, or public dissemination of knowledge.
  • Perform additional job-related duties as assigned.

Requirements

  • Master’s degree in Neuroscience, Psychology, Engineering, Computer Science or equivalent combination of education and experience is required.
  • 5+ years of experience in data analysis and data science fundamentals (e.g., algorithms, data structures, data visualization, machine learning), preferably in a clinical or research setting.
  • 5+ years of experience in at least one scientific programming language (e.g., Python/R, Matlab) and related toolboxes or frameworks (e.g., Tidyverse, Scipy, Sklearn, Polars, Pytorch) is required.
  • 5+ years of experience working in a Linux environment, using version control systems (e.g., GitHub), and software virtualization platforms (e.g., Docker).
  • 5+ years of practical experience in Extract, Transform, Load (ETL) processes and database management languages (SQL,NoSQL), and familiarity with associated cloud computing services and frameworks (AWS, Azure, Terraform).

Benefits & conditions

119000.00 To 150000.00 (USD) Annually medical insurance, parental leave, 401(k) United States, New York, New York 215 East 50th Street (Show on map) Jun 05, 2026

About Child Mind Institute

We’re dedicated to transforming the lives of children and families struggling with mental health and learning disorders by giving them the help they need. We’ve become the leading independent nonprofit in children’s mental health by providing gold-standard evidence-based care, delivering educational resources to millions of families each year, training educators in underserved communities, and developing tomorrow’s breakthrough treatments., Our great compensation package and benefits include medical insurance, 401(k), paid parental leave, dependent care, discounted tickets and entertainment perks programs. For more information about our benefits, please visit our employee benefits website.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on diversityjobs.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

4:54 min

Development history of scientific computation libraries and PyViz tools

Radovan Kavický · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all