> Markdown version of [/jobs/ext/2566500-senior-data-engineer](https://www.wearedevelopers.com/jobs/ext/2566500-senior-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Engineer - **Company:** Amaris - **Location:** Boston, MA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon S3, Data Analysis, Bash Shell, Batch Processing, Big Data, Clinical Data Repository, Databases, System Configuration, Data Validation, Data Transmissions, Information Engineering, Data Integrity, Extract Transform Load (ETL), Data Transformation, Python (Programming Language), Machine Learning, SAS (Software), Batch Scripting, Extensible Markup Language (XML), Jupyter Notebook, Data Processing, Electronic Medical Records, Semi-structured Data, Influxdb, Data Analytics, Machine Learning Operations, Data Pipelines - **Published:** August 6, 2026 - **Apply:** https://www.disabledperson.com/jobs/74023779-senior-data-engineer ## About the Role * 7-10 years of experience in Data Engineering, Data Analytics, or a related technical field. * Strong proficiency in Python for data processing and automation. * Experience with AWS services, particularly Amazon S3, Athena, and SageMaker. * Experience working with Jupyter Notebooks. * Solid understanding of ETL processes, data transformation, and data validation. * Experience handling structured and semi-structured data formats, including CSV and XML. * Familiarity with batch scripting using Bash. * Strong analytical and troubleshooting skills with a focus on data quality and integrity. * Experience working with SQL and large datasets. * Excellent communication and collaboration skills in cross-functional environments. Nice to Have * Experience with InfluxDB or other time-series databases. * Knowledge of AI/ML workflows and predictive analytics. * Experience working with physiological, medical device, or EMR data. * Familiarity with SAS and statistical validation processes. * Experience in the healthcare, medical technology, or life sciences industry. * Exposure to biomedical data analysis or clinical data environments. ## Description We are looking for a Senior Data Engineer to join a multidisciplinary team focused on healthcare data engineering, analytics, and AI-driven solutions. In this role, you will work with large-scale physiological, medical device, and electronic medical record (EMR) datasets to support data quality, predictive modeling, and advanced analytics initiatives. You will collaborate with data scientists, software engineers, biomedical specialists, and IT teams to ensure reliable data pipelines, investigate data integrity issues, and contribute to the development and validation of machine learning models., * Retrieve, explore, and analyze data stored in Amazon S3 using AWS Athena and Amazon SageMaker. * Perform data wrangling by identifying missing data, time synchronization issues, anomalies, and measurement artifacts. * Transform and parse CSV and XML datasets into scalable databases such as InfluxDB. * Develop and maintain batch processing workflows using Python and Bash to inventory, process, and validate incoming device data. * Investigate data quality issues through Athena, SageMaker, and Jupyter Notebooks to identify problems related to data curation, system configuration, storage, or performance. * Troubleshoot data transfer and collection issues in collaboration with biomedical teams, IT stakeholders, and internal engineering teams. * Analyze physiological, vital sign, and EMR datasets to support predictive analytics and AI/ML model development. * Validate statistical and machine learning outputs by comparing Python-generated results with SAS analyses produced by biostatistics teams. * Simulate retrospective healthcare datasets using a Virtual Hospital Simulator to evaluate and test surveillance algorithms. * Contribute to continuous improvements in data engineering processes, data quality, and analytical workflows. ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [WeAreDevelopers LIVE - CSS is DOOMed](https://www.wearedevelopers.com/videos/1838-wearedevelopers-live-css-is-doomed) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Hate organising your photos? Try it with 5 Terabytes](https://www.wearedevelopers.com/videos/79-hate-organising-your-photos-try-it-with-5-terabytes) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)