Sr. Data Engineer

Horizontal Talent
Boston, MA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Working hours
Regular working hours

Tech stack

Adobe InDesign Artificial Intelligence Airflow Amazon Web Services Amazon S3 Computer Programming Information Engineering Extract Transform Load (ETL) Python (Programming Language) Workflow Management Systems Data Processing Cloud Platform System
+10 more
Real Time Systems Large Language Models Apache Spark Kubernetes Information Technology AWS Data Analytics Stream Processing Data Pipelines Docker Databricks

Job description

We are seeking a Senior Data Engineer who will play a pivotal role in shaping our data platform team. This position offers the opportunity to work with cutting-edge technology while making a meaningful impact on our data-driven initiatives. Responsibilities

  • Design and implement robust data pipelines utilizing Databricks and AWS services for both batch and real-time processing.
  • Architect and maintain a scalable lakehouse infrastructure, ensuring data is reliable and accessible for analytics and AI applications.
  • Collaborate with data scientists to develop feature store architecture and support predictive modeling efforts.
  • Build and manage ETL/ELT orchestration and stream processing pipelines, integrating diverse data sources.
  • Ensure compliance with HIPAA regulations across all data handling processes.
  • Mentor junior engineers and enforce best engineering practices within the team.
  • Participate in design discussions and contribute to special projects as needed.

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Data Engineering, or a related field.
  • 7-10 years of professional experience in data engineering, with a strong focus on cloud-based architectures.
  • Proficiency in Apache Spark and hands-on experience with Databricks.
  • Expertise in AWS data services including S3, Redshift, and Glue.
  • Strong programming skills in Python, with additional knowledge of SQL.

Preferred Skills

  • Experience with ETL orchestration tools such as Airflow and dbt.
  • Familiarity with container orchestration using Docker and Kubernetes.
  • Knowledge of vector databases and LLM use cases.
  • Excellent communication skills, with the ability to convey complex technical concepts to non-technical stakeholders.

We are committed to fostering a diverse, equitable, and inclusive workplace where all individuals feel valued and respected. We encourage candidates from all backgrounds to apply and contribute to our mission.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on disabledperson.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

3:43 min

The enduring legacy of the amazon S3 storage API

Chris Heilmann +3 · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

5:08 min

Automating data collection and managing crowdsourced training image sets

Kris Howard · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · WWC 2024

Videos

See all

Related articles

See all