> Markdown version of [/jobs/ext/1137778-data-engineer](https://www.wearedevelopers.com/jobs/ext/1137778-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Agile Resources - **Location:** Orlando, FL, United States (Remote available) - **Experience:** Expert - **Salary:** $124,800.0 - $145,600.0 - **Contract:** Temporary to permanent - **Skills:** Cloud Computing, Continuous Delivery, Continuous Integration, Information Engineering, Python (Programming Language), SQL Databases, Large Language Models, Apache Spark, Event Driven Architecture, Apache Kafka, Machine Learning Operations - **Published:** July 2, 2026 - **Apply:** https://jobs.gotoagile.com/index.smpl?arg=jb_apply&POST_ID=14147558 ## About the Role * 7+ years of sophisticated data engineering experience, highlighted by mastery-level knowledge of Spark tuning, partitioning, and cloud infrastructure. * Extensive hands-on experience designing secure, ACID-compliant storage layers, ideally utilizing modern unified cataloging tools. * Expertise in Python and SQL, coupled with practical experience supporting ML lifecycle management (such as tracking, feature stores, or LLM integrations). * A proven track record of integrating disparate, complex data sources and maintaining high-availability production environments at enterprise-scale. * Familiarity with event-driven architectures (like Kafka), continuous integration/continuous deployment (CI/CD) pipelines, and infrastructure-as-code will be a plus. ## Description We are partnering with an industry leader in the construction supplies and equipment space to find a Sr. Data Engineer to spearhead the evolution of their global, next-generation data ecosystem. If you thrive on solving intricate, massive-scale data puzzles, optimizing bleeding-edge distributed computing environments, and laying the groundwork for sophisticated machine learning and LLM operations, this is your next definitive career move. In this role, you will act as the architect behind a mission-critical platform, transforming raw datasets into powerful, secure, and production-ready intelligence that fuels executive-level decisions. You will enjoy a high degree of technical ownership, bridging the gap between advanced cloud engineering and real-world AI applications. Here’s what you’ll be doing: * Architect and optimize high-throughput, distributed computing frameworks and modern Lakehouse structures to handle massive data velocity and volume. * Construct robust, production-grade features and pipelines that seamlessly operationalize large language models (LLMs) and predictive machine learning models. * Build automated ingestion frameworks across multi-cloud environments, utilizing both real-time streaming and scheduled batch processing. * Establish solid data quality, unified cataloging, and access controls while aggressively optimizing cluster performance and cloud infrastructure costs. * Collaborate with cross-functional leadership, finance, and operations teams to convert complex strategic goals into highly scalable technical solutions. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [How to Benchmark Your Apache Kafka](https://www.wearedevelopers.com/videos/76-how-to-benchmark-your-apache-kafka) - [Fault Tolerance and Consistency at Scale: Harnessing the Power of Distributed SQL Databases](https://www.wearedevelopers.com/videos/1146-fault-tolerance-and-consistency-at-scale-harnessing-the-power-of-distributed-sql-databases) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Implementing continuous delivery in a data processing pipeline](https://www.wearedevelopers.com/videos/73-implementing-continuous-delivery-in-a-data-processing-pipeline) - [TiDB, One Layer at a Time: How Distributed SQL Became an Agentic AI Backbone](https://www.wearedevelopers.com/videos/100117-tidb-one-layer-at-a-time-how-distributed-sql-became-an-agentic-ai-backbone) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)