> Markdown version of [/jobs/ext/3428702-data-engineer](https://www.wearedevelopers.com/jobs/ext/3428702-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Mass General Brigham - **Location:** Somerville, MA, United States (Remote available) - **Experience:** Experienced - **Salary:** $75,275.0 - $109,554.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Microsoft Azure, Clinical Data Repository, Computer Programming, Continuous Integration, Information Engineering, Extract Transform Load (ETL), Data Transformation, Data Systems, Data Warehousing, Relational Databases, Python (Programming Language), Machine Learning, Metadata, Reference Data, Cloud Services, Azure Data Lake, SQL Databases, Data Streaming, Tableau (Software), Scripting, Cloud Platform System, Azure Data Factory, Snowflake, Apache Spark, Microsoft Fabric, Data Lakes, Pyspark, Information Technology, Apache Kafka, Stream Analytics, Data Pipelines - **Published:** September 9, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=8f5f05e5a352797a ## About the Role * Bachelor's degree in Computer Science, a related field, or equivalent experience. * 2-5 years of experience in data engineering, data warehousing, data pipeline development, or a related technical role. * Hands-on experience with SQL and at least one programming or scripting language such as Python, PySpark, or Spark notebooks. * Experience developing data pipelines, ETL/ELT processes, and data transformation workflows. * Working knowledge of cloud data platforms such as Azure, AWS, or GCP; Azure Data Lake, Azure Data Factory, Microsoft Fabric, or Snowflake experience is strongly preferred. * Experience working with relational databases, data warehouses, or large-scale reporting environments. * Familiarity with data modeling, metadata, data quality, workload management, and dependency management concepts. * Ability to collaborate with cross-functional teams and communicate technical information clearly to technical and non-technical partners. Preferred Experience * Experience with Microsoft Fabric, Fabric Spark notebooks, Azure DevOps, CI/CD, or dbt. * Experience with Snowflake features such as Snowpipe, SnowSQL, Snowsight, or Data Streams. * Experience with real-time analytics technologies such as Spark, Kafka, or Event Hub. * Healthcare data experience, including clinical data, Epic, Clarity, payer data, or reference data. * Experience with business intelligence tools such as Tableau or similar platforms. Skills That Will Help You Succeed * Strong analytical, troubleshooting, and problem-solving skills. * Ability to design reliable, scalable, and maintainable data solutions. * Comfort working independently while staying connected to team priorities. * Strong communication skills and the ability to collaborate with partners across technical and business teams. * Ability to prioritize work, manage multiple initiatives, and adapt in a dynamic environment. * Interest in continuous learning and applying new data technologies, tools, and best practices. * Commitment to teamwork, respect, inclusion, and high-quality service. ## Description Mass General Brigham Digital is transforming health care through data, analytics, and digital innovation. We are investing in enterprise data platforms, advanced analytics, artificial intelligence, and machine learning to improve patient care, support research, and make insights easier to use across the system. We are seeking a self-motivated Data Engineer to join our Data Lake Engineering team. In this role, you will design, build, and maintain scalable data pipelines and cloud-based data platforms that help teams across Mass General Brigham access reliable, high-quality data. This is a great opportunity for someone who enjoys solving complex data challenges, optimizing data systems, and partnering with cross-functional teams to support meaningful work in health care. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [A Data Mesh needs Open Metadata](https://www.wearedevelopers.com/videos/505-a-data-mesh-needs-open-metadata) - [How Cisco embraced a DevOps culture within its network engineering team](https://www.wearedevelopers.com/videos/99-how-cisco-embraced-a-devops-culture-within-its-network-engineering-team) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)