> Markdown version of [/jobs/ext/1857154-lead-data-engineer-pipelines-spark-streaming-and-spark-offline](https://www.wearedevelopers.com/jobs/ext/1857154-lead-data-engineer-pipelines-spark-streaming-and-spark-offline). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead Data Engineer - Pipelines, Spark Streaming and Spark Offline - **Company:** JPMorgan Chase & Co. - **Location:** Tampa, FL, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Business Analytics Applications, Big Data, Cloud Engineering, Computer Programming, Data as a Services, Information Engineering, Extract Transform Load (ETL), Data Security, Data Systems, Data Warehousing, Distributed Systems, Apache Hadoop, Information Lifecycle Management, Python (Programming Language), NoSQL, Software Systems, Data Processing, Feature Engineering, Flask (Web Framework), Apache Spark, Data Lakes, Pyspark, Kubernetes, Apache Flink, Integration Frameworks, Apache Kafka, Spark Streaming, Data Management, Stream Processing, Data Pipelines, AWS EKS, Docker, Databricks - **Published:** July 4, 2026 - **Apply:** https://www.jobmonkeyjobs.com/career/27824151/Lead-Data-Engineer-Pipelines-Spark-Streaming-Spark-Offline-Florida-Tampa-7463 ## About the Role * Formal training or certification on Data Engineering concepts and 5+ years applied experience * Demonstrated experience using enterprise-authorized AI capabilities within the work environment to support data engineering workflows with strong validation habits and awareness of data sensitivity * Ability to review and validate AI-assisted outputs (e.g., model/design summaries or operational checklists) before use, escalating when uncertain and following data handling requirements * Experienced programming skills with Python, PySpark * Experience across the data lifecycle, building Data frameworks, working with Data lakes * Experience with Batch and Real time Data processing with Spark or Flink and Batch and Real time feature engineering with Spark or Flink or data brick * Working knowledge of AWS Glue and EMR usage for Data processing and real time data processing and features using Flink or Data brick live tables or Spark streaming * Experience working with Databricks and data brick live tables * Experience working in building services using Glue, Lamida, EMR or Flask, and deploying them on AWS EKS or Kubernetes * Working experience with both relational and NoSQL databases * Experience in ETL data pipelines both batch and real-time data processing, Data warehousing, NoSQL DB Preferred qualifications, capabilities, and skills * Expertise in Amazon Web Services (AWS), Docker, and Kubernetes for cloud-native and containerized data solutions * Experience in big data technologies: Hadoop, Spark, Kafka, Flink * Experience in distributed system design and development ## Description As a Lead Data Engineer at JPMorganChase within the Commercial & Investment Bank, you are an integral part of an agile team that works to enhance, build, and deliver data collection, storage, access, and analytics solutions in a secure, stable, and scalable way. As a core technical contributor, you are responsible for maintaining critical data pipelines and architectures across multiple technical areas within various business functions in support of the firm's business objectives., * Collaborate with all of JPMorgan's lines of business and functions to delivery software solutions * Experiment, Architect, develop and productionize efficient Data pipelines, Data services and Data platforms contributing to the business * Design and implement highly scalable, efficient and reliable data processing pipelines and perform analysis and insights to drive and optimize business result * Design and develop features and entities for ML and rule using spark or any bigdata environment * Acts on previously identified opportunities to converge physical, IT, and data security architecture to manage access * Applies reuse-first, AI-assisted practices within delivery and operational routines (e.g., backup/recovery validation and access control review support), ensuring traceability/auditability and alignment to resiliency and security expectations ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Hosting a modern justice system](https://www.wearedevelopers.com/videos/332-hosting-a-modern-justice-system) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers)