> Markdown version of [/jobs/ext/519028-data-engineer](https://www.wearedevelopers.com/jobs/ext/519028-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Child, Inc. - **Location:** New York, NY, United States - **Experience:** Expert - **Salary:** $119,000.0 - $150,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Data Analysis, Microsoft Azure, Cloud Computing, Data Infrastructure, Extract Transform Load (ETL), Data Transformation, Data Security, Data Structures, Data Visualization, Linux, Github, Python (Programming Language), MATLAB, Machine Learning, NoSQL, SciPy, SQL Databases, Virtualization Technology, Pytorch, Model Validation, Containerization, Scikit Learn, Kubernetes, Information Technology, Data Analytics, Terraform, Software Version Control, Data Pipelines, Docker - **Published:** June 6, 2026 - **Apply:** https://diversityjobs.com/career/15877285/Data-Engineer-New-York-New-York ## About the Role * Master's degree in Neuroscience, Psychology, Engineering, Computer Science or equivalent combination of education and experience is required. * 5+ years of experience in data analysis and data science fundamentals (e.g., algorithms, data structures, data visualization, machine learning), preferably in a clinical or research setting. * 5+ years of experience in at least one scientific programming language (e.g., Python/R, Matlab) and related toolboxes or frameworks (e.g., Tidyverse, Scipy, Sklearn, Polars, Pytorch) is required. * 5+ years of experience working in a Linux environment, using version control systems (e.g., GitHub), and software virtualization platforms (e.g., Docker). * 5+ years of practical experience in Extract, Transform, Load (ETL) processes and database management languages (SQL,NoSQL), and familiarity with associated cloud computing services and frameworks (AWS, Azure, Terraform). ## Description As part of the Center for Data Analytics, Innovation, and Rigor team, you will report to Rubric Engineering and Measurement Specialist. You will develop infrastructure to support large-scale AI evaluation frameworks. You will design scalable data pipelines for generating and processing synthetic data, implement secure data storage solutions, and create infrastructure for real-time model evaluation and monitoring. You will use common frameworks, platforms, and languages, such as Python, SQL, GitHub, containerization tools (e.g., Docker, Kubernetes), and cloud computing infrastructures (e.g., AWS, Azure) to build robust and scalable data infrastructure that support our AI research initiatives. This is an exempt, full-time, hybrid position located in our NYC headquarters office or other relevant location. This position requires a minimum of four (4) days per week in the office, on a schedule determined by your supervisor. The in-office requirement and schedule are subject to change based on the needs of the program and the organization. You Will: * Create and maintain scalable data pipelines for efficient storage and retrieval of multimodal data, with particular emphasis on clinical, natural language, and multi-turn response data. * Create pipelines for data transformation, preprocessing, and management. * Ensure data quality, security, and compliance with privacy regulations for handling sensitive data. * Perform quality assurance of pipelines/processes to maintain integrity throughout the data lifecycle. * Create interactive visualizations and dashboards to communicate data insights and pipeline performance metrics. * Write documentation and relevant text for scientific, clinical, or public dissemination of knowledge. * Perform additional job-related duties as assigned. ## Related Videos - [Python Data Visualization @ Deepnote (w/ PyViz overview)](https://www.wearedevelopers.com/videos/113-python-data-visualization-deepnote-w-pyviz-overview) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story)