Senior Data Engineer, Data Science Infrastructure

Liberty Mutual Insurance Company
Boston, MA, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Compensation
$83,000.0 - $157,000.0
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Artificial Intelligence Amazon Web Services Cloud Computing Code Generation Code Review Computer Engineering Data Discovery Information Engineering Data Infrastructure Data Integration Extract Transform Load (ETL)
+23 more
Data Mining Data Systems Data Warehousing Programming Tools Python (Programming Language) Operational Databases Scrum Methodology Query Optimization Cloud Services SQL Databases Data Streaming Parquet Feature Engineering Snowflake Prompt Engineering Apache Spark Pyspark Information Technology Low Latency Data Management Machine Learning Operations Data Pipelines Databricks

Job description

Modeling Data Solutions is seeking experienced data engineers to join our Data Solutions job family. This is an exciting opportunity to join the US Data Science Infrastructure department helping to support creating cutting edge pricing programs. You will play a critical role in designing and developing the data solutions needed for research and development as well as providing front line data support in launching new products into market., * Lead the design, development, and maintenance of scalable and efficient data pipelines and ETL/ELT processes to support analytics and pricing models.

  • Architect and implement high performance data integration solutions using modern tools and cloud services (AWS preferred) with attention to latency, throughput and cost.
  • Implement and maintain automated data quality and testing frameworks (unit, integration, regression, anomaly detection) to ensure data correctness and trust for downstream analytics
  • Collaborate with cross functional teams, including data scientists, product analysts, software engineers, to gather requirements, design solutions, and provide technical guidance for production launches
  • Optimize and tune data infrastructure for performance, scalability, reliability and cost-efficiency (query tuning, partitioning, resource configuration, storage formats such as Parquet/Delta).
  • Mentor and guide junior data engineers, promoting engineering best practices, code review and thoughtful design.
  • Stay up-to-date with advancements in data engineering, evaluate new tools and recommend adoption where appropriate.
  • Identify opportunities to apply Gen AI for data discovery, lineage summarization, automated documentation, query generation and developer productivity (prompt engineering and code generation)
  • Incorporate Gen AI capabilities into data workflows and developer tooling and help operationalize safe, cost-effective Gen AI integrations.

Requirements

  • Have strong technical aptitude with data extraction and data engineering platforms (Python, Spark/PySpark, SQL, cloud computing: AWS and Snowflake).
  • Demonstrate 5+ years of experience building production data pipelines and platforms, with both batch and streaming experience
  • Possess hands-on experience with data quality testing frameworks and practices, and experience implementing automated data tests and validation.
  • Have practical experience collaborating in Agile teams and applying Agile best practices (sprint planning, refinement, retrospective).
  • Have excellent communication skills and the ability to work with technical and non-technical stakeholders.

Preferred qualifications:

  • Familiarity with ML pipelines and supporting feature stores or feature engineering workflows (Databricks).
  • Prior mentoring or technical leadership experience., * Strong written and oral communication skills required
  • Bachelor`s Degree in Computer Science, Computer Engineering, or related discipline preferred
  • Master`s in same or related disciplines strongly preferred
  • 3-5 years experience in coding for data management, data warehousing, or other data environments, including, but not limited to, working in Python, SQL, ETL, Spark, Snowflake.

About the company

Pay Philosophy: The typical starting salary range for this role is determined by a number of factors including skills, experience, education, certifications and location. The full salary range for this role reflects the competitive labor market value for all employees in these positions across the national market and provides an opportunity to progress as employees grow and develop within the role. Some roles at Liberty Mutual have a corresponding compensation plan which may include commission and/or bonus earnings at rates that vary based on multiple factors set forth in the compensation plan for the role.

At Liberty Mutual, our goal is to create a workplace where everyone feels valued, supported, and can thrive. We build an environment that welcomes a wide range of perspectives and experiences, with inclusion embedded in every aspect of our culture and reflected in everyday interactions. This comes to life through comprehensive benefits, workplace flexibility, professional development opportunities, and a host of opportunities provided through our Employee Resource Groups. Each employee plays a role in creating our inclusive culture, which supports every individual to do their best work. Together, we cultivate a community where everyone can make a meaningful impact for our business, our customers, and the communities we serve.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:50 min

How Parquet metadata enables efficient data reading

Matthias Niehoff Matthias Niehoff · WWC Europe 2026

1:33 min

Integrating internal APIs and maintaining data sovereignty

Mahran Meißner Mahran Meißner · WWC Europe 2026

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all