> Markdown version of [/jobs/ext/623790-vice-president-data-pipeline-engineer](https://www.wearedevelopers.com/jobs/ext/623790-vice-president-data-pipeline-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Vice President, Data Pipeline Engineer - **Company:** LIFE AT BALANCE - **Location:** Lake Mary, FL, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, Data Analysis, Information Engineering, Data Infrastructure, Data Integration, Extract Transform Load (ETL), Data Transformation, Dataspaces, Distributed Systems, Graph Database, Intrusion Detection Systems, Python (Programming Language), Performance Tuning, Reference Data, SQL Databases, Data Streaming, Unstructured Data, Workflow Management Systems, Data Logging, Data Processing, Snowflake, Apache Spark, Usage Tracking, Event Driven Architecture, Build Management, Semi-structured Data, Api Design, Data Pipelines, Databricks - **Published:** June 23, 2026 - **Apply:** https://www.dice.com/job-detail/0a739f28-f638-4657-a87e-0ec33c4a58da ## About the Role * Bachelor's degree in a related discipline or equivalent work experience required. An advanced degree with a preference in statistics/statistical analysis is preferred. * At least six years' total work experience, with at least 3 years' experience with a strong focus on data analysis and business intelligence is preferred. * Extensive experience in data engineering, building and scaling production-grade data pipelines. * Deep hands-on expertise in Python, Spark, and SQL, with strong experience in ETL/ELT frameworks and orchestration tools. * Proven ability to design and operate high-volume, resilient pipelines across batch, streaming, and distributed environments. * Strong understanding of structured and semi-structured data modeling, including time-series and event-driven architectures. * Experience designing data transformation and normalization layers, including schema evolution and backward compatibility. * Expertise with modern data platforms (e.g., Snowflake, AWS, Databricks), lakehouse architectures, and API-based data integration. * Strong capabilities in performance tuning, cost optimization, and implementing data quality, monitoring, logging, and lineage frameworks. * Domain experience with financial datasets (market data, pricing, reference data, portfolio holdings, transactions, corporate actions) and familiarity with key vendors (e.g., Bloomberg, ICE, MSCI). * Exposure to knowledge graph/ontology-driven systems, entity resolution workflows, AI/LLM-based unstructured data integration (e.g., documents, PDFs), and data entitlements, licensing, and usage tracking is preferred. ## Description We're seeking a future team member for the role of Data Pipeline Engineer to join our Data Innovation team. In this role, you will design and build the data pipelines that power our Investment Data Standard (IDS) and knowledge graph, enabling a unified, high-quality data ecosystem that supports analytics, AI, and client-facing solutions. You will partner closely with the Ontology/Knowledge Architecture lead and collaborate across platform, product, and data teams to deliver scalable, production-ready solutions aligned to our broader data transformation strategy. This role is located in Pittsburgh, PA or Lake Mary, FL. In this role, you'll make an impact in the following ways: * Design and build scalable pipelines to ingest and process data from internal platforms and external vendors across batch, streaming, and near real-time patterns. * Transform diverse data formats (APIs, flat files, streaming, unstructured) into clean, standardized time-series and event-driven datasets aligned to IDS entity models. * Develop reusable frameworks to normalize identifiers, symbology, units, hierarchies, and event data (e.g., corporate actions, transactions). * Partner with Ontology/Knowledge architecture team to map source data to canonical entities, relationships, and attributes, enabling graph ingestion and entity resolution. * Implement robust data quality controls (completeness, accuracy, consistency, schema drift, anomaly detection) with full lineage, provenance, and traceability (source IDS product). * Enable multi-vendor data ingestion, comparison, and reconciliation, including source prioritization, hierarchy logic, and coverage/quality analytics. * Build modular, reusable, cloud-native pipelines optimized for scale, performance, and cost (e.g., Snowflake), with monitoring and SLA-driven reliability. * Collaborate cross-functionally to translate business and data requirements into production-ready pipelines and support downstream distribution via APIs, data products, and client platforms. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [API Design - Getting Started](https://www.wearedevelopers.com/videos/33-api-design-getting-started) - [Cutting LLM Costs Without Cutting Quality: How to Beat Proprietary LLMs with Fine-Tuned Open Source](https://www.wearedevelopers.com/videos/100151-cutting-llm-costs-without-cutting-quality-how-to-beat-proprietary-llms-with-fine-tuned-open-source) - [How Cisco embraced a DevOps culture within its network engineering team](https://www.wearedevelopers.com/videos/99-how-cisco-embraced-a-devops-culture-within-its-network-engineering-team) - [OLTP in the Lakehouse: Redefining Data for AI Workloads](https://www.wearedevelopers.com/videos/2038-oltp-in-the-lakehouse-redefining-data-for-ai-workloads) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk)