> Markdown version of [/jobs/ext/3037617-data-engineer-lead](https://www.wearedevelopers.com/jobs/ext/3037617-data-engineer-lead). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer Lead - **Company:** State Street - **Location:** Boston, MA, United States (Remote available) - **Experience:** Expert - **Salary:** $120,000.0 - $202,500.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Agile Methodology, Data Analysis, Microsoft Azure, Big Data, Cloud Computing, Configuration Management, Cyber Security, Information Systems, Computer Programming, Databases, System Configuration, Information Engineering, Data Governance, Data Infrastructure, Data Integrity, Extract Transform Load (ETL), Data Security, Data Structures, Data Systems, Data Warehousing, DevOps, Dimensional Modeling, Document-Oriented Databases, Microsoft SQL Server, SQL Azure, Oracle (Applications), Query Optimization, Raw Data, Standard Sql, Server Administration, Software Engineering, SQL Databases, Data Streaming, Unstructured Data, Data Import/Export, Data Processing, Data Storage Technologies, Data Ingestion, Azure Data Factory, Snowflake, Database Optimization, Git, Data Lakes, Kubernetes, Information Technology, Non-relational Database, Data Management, Azure Synapse Analytics, Software Version Control, Data Pipelines, Docker, Databricks, Programming Languages - **Published:** September 23, 2026 - **Apply:** https://www.careerbuilder.com/job-details/sr-data-engineering-lead-onsite-vp-boston-ma--5bf9afc1-d49c-42b4-bd11-460f5e1144a7 ## About the Role * Hands-on experience with Azure services such as Azure Synapse Analytics, Azure Data Factory, Azure Databricks, Azure SQL Database, etc. * Strong experience in data engineering, including data pipeline development, ETL/ELT processes, and data modeling. * Strong proficiency in SQL and experience with programming languages such as Scala, or Java * In-depth knowledge of SQL and experience with relational and non-relational databases such as Snowflake, SQLServer, Oracle * Knowledge of data warehousing concepts, dimensional modeling, and best practices * Experience working in a multi-developer environment, using version control (i.e. Git) * Excellent problem-solving skills and ability to work independently as well as part of a team * Azure certifications such as Azure Data Engineer Associate or Azure Solutions Architect is a plus * Excellent communication skills and the ability to collaborate effectively with cross-functional teams, * Bachelor's Degree level qualification in a computer or IT related subject * 10+ years of overall IT industry experience * 8+ years of overall Bigdata data pipeline experience * 8+ years of experience as a Data Engineer, with a focus on designing and implementing data solutions on the Azure Databricks * 8+ years of experience on cloud-based development including Azure Services, Azure Devops, Kubernetes, Docker * Experience on Snowflake is plus Additional requirements * Communicate effectively in a professional manner both written and orally * Team player with a positive attitude enthusiasm initiative and self-motivation * Ability to multi-task energetic fast learner & problem solver * Experience of working in the financial industry * Experience with agile development methodology, Agile Programming Methodologies, Analysis Skills, Best Practices, Cloud Applications, Cloud Computing, Communication Skills, Compensation and Benefits, Computer Programming, Computer Security, Cross-Functional, Data Analysis, Data Import/Export, Data Management, Data Modeling, Data Processing, Data Quality, Data Science, Data Storage, Data Structures, Data Warehousing, Database Administration, Database Extract Transform and Load (ETL), Database Optimization, Database Technology, DevOps, Dimensional Modeling, Docker, Genetics, Git, Identify Issues, Incentive Programs, Information Technology & Information Systems, Information/Data Security (InfoSec), Java, Maintain Compliance, Microsoft Windows Azure, Military, Monitor Regulations, Performance Management, Privacy Regulations, Problem Solving Skills, Process Development, Process Flow, Process Modeling, Profit & Loss, Programming Languages, Query Optimization, Regulatory Compliance, Risk Management, SQL (Structured Query Language), Sales, Scala Programming Language, Software Administration, Software Engineering, Source Code/Configuration Management (SCM), Structured Data, Systems Administration/Management, Systems Scalability, Tax Planning, Team Player, Unstructured Data ## Description As a Data Engineer Lead, you will be responsible for developing, maintaining, and optimizing data pipelines and systems that support the acquisition, storage, transformation, and analysis of large volumes of data. You will collaborate with cross-functional teams, including analysts, software engineers, and operations teams to ensure the availability, reliability, and integrity of data for various business needs. This role requires strong technical expertise in data engineering principles, database management, and programming skills. We are looking for candidate with good knowledge on Bigdata technology and strong development experience with Databricks. You will help to migrate existing on-prem applications to cloud (Azure preferred), create and maintain new data applications relying on experience and judgment to plan and accomplish goals, while working with other globally situated team members. What you will be responsible for As Data Engineer Lead you will * Design, develop, and maintain data pipelines in Azure for ingesting, transforming, and loading data from various sources into centralized Azure data lakes and Databricks Delta Lake. Ensure data quality and integrity throughout the process. * Implement efficient ELT/ETL processes to ensure data quality, consistency, and reliability. Develop transformation processes to clean, aggregate, and enrich raw data, ensuring it is in the appropriate format for downstream analysis and consumption. Integrate data from diverse sources to provide a unified view of information. * Design and implement efficient data models and database schemas that support the storage and retrieval of structured and unstructured data. Optimize data storage and access for performance and scalability. * Implement knowledge of modern data processing principles to streamline data import/transformation processes. Leverage modern data pipeline tools to reduce human attention during ETL process. Ensure the efficiency and reliability of data ingestion and processing. * Work closely with cross-functional teams to understand data requirements and translate them into technical solutions. * Monitor data pipelines, troubleshoot issues, and ensure data integrity and security. * Implement data quality controls and validation processes to identify and rectify data anomalies, inconsistencies, and errors. Collaborate with stakeholders to define and enforce data governance standards and policies. * Identify performance bottlenecks in data pipelines and database systems and optimize queries, data structures, and infrastructure configurations to improve overall system performance and scalability. * Implement appropriate security measures to protect sensitive data and ensure compliance with data privacy regulations. Monitor and address data security vulnerabilities and risks. * Collaborate with cross-functional teams, including data scientists, analysts, and software engineers, to understand data requirements and deliver effective data solutions. Document data engineering processes, data flows, and system configurations. * Stay updated with the latest trends, tools, and technologies in the field of data engineering. Proactively identify opportunities to improve data engineering practices and contribute to the evolution of data infrastructure. * Provide support to the development team with managing multiple instances of databases and servers, implementing complex queries with proper tuning, provide input to design impacting data, manage data infrastructure in Azure (DataLake, DataWarehouse and Synapse). ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)