Data Scientist Lead, Vice President

JPMorgan Chase & Co.
Plano, TX, United States
1 day ago
Apply on www.themuse.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Airflow Code Review Information Engineering Extract Transform Load (ETL) Data Structures Database Queries Distributed Computing Environment Python (Programming Language) Object-Oriented Software Development Operational Databases Performance Tuning
+14 more
Systems Development Life Cycle Cloud Services Runbook SQL Databases Data Streaming Strategies of Testing Snowflake Apache Spark Data Layers Pyspark Information Technology Software Version Control Data Pipelines Databricks

Job description

We are seeking a Data Engineering Lead to help build and evolve a high-quality measurement data foundation that enables trusted analytics and decision-making at scale. This role focuses on designing and delivering resilient datasets, pipelines, and reusable metrics that support hypothesis-driven analyses and experiments across the product development lifecycle (PDLC).

You’ll be hands-on where needed, drive engineering standards, and help teams move faster by improving reliability, observability, and usability across the data lifecycle-so leaders can clearly see what’s driving value, what’s creating friction, and what operating-model shifts materially improve outcomes as teams become more agentic., * Design, build, and operate scalable data pipelines (batch and/or streaming) with clear SLAs, monitoring, and incident response practices.

  • Develop and curate trusted data products (e.g., conformed dimensions, event models, marts) with strong documentation and clear ownership.
  • Build and maintain well-defined metrics and feature-ready datasets that enable measurement of AI adoption and productivity outcomes (e.g., reusable aggregates, cohorting, time-windowed measures), including change control as definitions evolve.
  • Drive data quality and governance through validations, reconciliations, lineage, access controls, retention, and auditability aligned to requirements.
  • Develop and operate workflow orchestration (e.g., Apache Airflow) to schedule, monitor, and manage data movement and transformations.
  • Model and transform data for analytics using SQL/dbt to support trusted reporting and repeatable measurement.
  • Write production-grade Python/PySpark with disciplined testing, performance tuning, and maintainable design.
  • Partner with analytics, product, and engineering stakeholders to define requirements, success criteria, and consistent interpretation of key measures-particularly where inputs span finance business cases, PDLC/SDLC tools, and AI tool logs.
  • Establish and enforce engineering best practices (version control, code review, testing strategy, deployment processes, runbooks) and continuously improve observability and cost/performance (freshness, completeness, timeliness, scalability, spend).
  • Mentor and develop a team of 2, influencing technical direction through standards, reviews, and knowledge sharing.

Requirements

  • Bachelor’s degree in Computer Science, Engineering, or equivalent practical experience.
  • 5+ years of hands-on experience delivering production data solutions in a fast-paced engineering environment (actively coding and owning outcomes).
  • Strong software engineering fundamentals (system design, data structures, object-oriented programming, testing strategies, and end-to-end development lifecycle).
  • Strong understanding of data modeling (conceptual, logical, physical), including dimensional, normalized, and event-based approaches.
  • Hands-on experience with Databricks and large-scale distributed data processing/performance tuning (Spark/PySpark).
  • Strong SQL skills and experience with modern transformation tooling (e.g., dbt), including building maintainable, testable data codebases.
  • Experience designing and operating orchestration pipelines using Airflow (or equivalent), including backfills, retries, and operational monitoring.
  • Demonstrated rigor building and maintaining trusted metrics (definitions, edge cases, validation/testing, documentation) and keeping them reliable as upstream sources change.
  • Demonstrated ability to lead delivery in complex environments with multiple stakeholders and ambiguous requirements.

Preferred Qualifications

  • Experience with modern lakehouse/warehouse patterns and broader cloud data platforms (e.g., Databricks, Snowflake).
  • Experience with BI/semantic layers and metrics management practices.
  • Exposure to experimentation or hypothesis-driven analytics approaches (e.g., measurement design to support tests, rollouts, and pre/post evaluation); deep causal specialization not required.

Benefits & conditions

We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process.

About the company

Chase is a leading financial services firm, helping nearly half of America’s households and small businesses achieve their financial goals through a broad range of financial products. Our mission is to create engaged, lifelong relationships and put our customers at the heart of everything we do. We also help small businesses, nonprofits and cities grow, delivering solutions to solve all their financial needs.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.themuse.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:50 min

Introduction and the value of runbooks

Hila Fish · World Congress 2023

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

Videos

See all

Related articles

See all