Data Engineer

AgileEngine, LLC
Boca Raton, FL, United States
2 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Regular working hours
Languages
English
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Airflow Amazon Web Services Data Analysis Big Data Databases Data Architecture Data Infrastructure Data Integrity Extract Transform Load (ETL) Data Retrieval
+23 more
Data Warehousing Document-Oriented Databases Elasticsearch Python (Programming Language) Machine Learning NoSQL Simple Data Format SQL Databases Parquet Data Processing Google Cloud Data Ingestion Snowflake Apache Spark Jupyter Pandas Build Management Data Lakes Pyspark Information Technology Avro Terraform Data Pipelines

Job description

We are looking for a Senior Data Engineer to design and build scalable data lakes, warehouses, and lakehouse architectures supporting a thematic research platform that processes large volumes of financial data daily. You will implement Python-based ETL/ELT pipelines, orchestrate workflows with Airflow, develop ingestion workflows from third-party APIs, and work with Snowflake, Spark, and AWS to deliver high-performance data infrastructure. The role combines hands-on engineering with technical consulting responsibilities, translating business goals into data architecture roadmaps.

WHAT YOU WILL DO

  • Design and implement Python Data Engineering solutions;

  • Design and build scalable Data Lakes, Data Warehouses, and Data Lakehouses;

  • Design and implement robust ETL/ELT processes at scale using Python, incorporating modern pipeline orchestration tools like Airflow;

  • Develop sophisticated ingestion workflows from diverse 3rd party APIs and data sources;

  • Manage and optimize various file formats (Parquet, Avro, ORC) and columnar storage to ensure high-performance data retrieval;

  • Work with AI development tools to support and accelerate ongoing development, machine learning initiatives and advanced analytics;

  • Act as a technical consultant for stakeholders and leadership to gather requirements, understand business goals, and translate them into technical roadmaps

Requirements

If you’re looking for a place to grow, make an impact, and work with people who care, we’d love to meet you!, You must be authorized to work for ANY employer in the US (e.g., s, TN visa holders, U4U with EAD), as we are unable to sponsor or take over employment visa sponsorship at this time;

  • Bachelor’s degree in computer science/engineering or other technical field, or equivalent experience;

  • 5+ years of experience with Python (strong, hands-on Python experience is a must);

  • 5+ years of experience with data processing, manipulation, and analytics libraries like Pandas, Polars, PySpark or DuckDB;

  • 2+ years of experience with Big Data technologies (Spark, Snowflake);

  • Expert-level knowledge of pipeline orchestration using Airflow or similar industry-standard tools;

  • Deep understanding of Medallion Architecture, columnar file formats, and diverse database technologies (SQL, NoSQL, and Lakehouse architectures);

  • Proven ability to work with 3rd party APIs for complex data ingestion tasks;

  • Proficiency with modern Cloud platforms (AWS, Google Cloud Platform, Snowflake) and advanced SQL optimization;

  • Exceptional soft skills with a proven ability to gather requirements from leadership and collaborate effectively across cross-functional teams;

  • Excellence in optimizing complex data pipelines and troubleshooting data latency or consistency issues in massive datasets;

  • A self-starter mindset, regularly investigating more efficient data architectures and AI development tools to improve pipeline performance;

  • Taking pride in data integrity and the accuracy of the end-to-end pipelines and architectures you build;

  • Strong communication skills for seamless global collaboration with stakeholders and distributed teams;

  • Upper-intermediate English level.

NICE TO HAVES

  • Familiarity with the fintech industry, understanding of financial data, regulatory requirements, and business processes specific to the domain;

  • Documentation skills to document data pipelines, architecture designs, and best practices for knowledge sharing and future reference;

  • OpenSearch, Elasticsearch;

  • AWS Sagemaker Studio, Jupyter for analyze data;

  • Terraform;

  • Scala.

Benefits & conditions

Competitive compensation: USD-based pay with education, fitness, and team activity budgets.

  • Exciting projects: Modern solutions with Fortune 500 and top product companies.

About the company

AgileEngine is an Inc. 5000 company that creates award-winning software for Fortune 500 brands and trailblazing startups across 17+ industries. We rank among the leaders in areas like application development and AI/ML, and our people-first culture has earned us multiple Best Place to Work awards.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · WWC 2024

3:02 min

Audience Q&A on data formats and engine tradeoffs

Matthias Niehoff Matthias Niehoff · WWC Europe 2026

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:58 min

Analyzing production code coverage data using pandas

Markus Harrer Markus Harrer · WWC 2021

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

Videos

See all

Related articles

See all