> Markdown version of [/jobs/ext/1339718-staff-ai-data-engineer](https://www.wearedevelopers.com/jobs/ext/1339718-staff-ai-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff AI Data Engineer - **Company:** CR FITNESS CAPE CORAL, LLC - **Location:** Cleveland, OH, United States - **Experience:** Expert - **Salary:** $125,000.0 - $145,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Big Data, Cloud Database, Computer Programming, Databases, Continuous Delivery, Continuous Integration, Data Validation, Information Engineering, Extract Transform Load (ETL), Data Warehousing, Apache Hadoop, Python (Programming Language), Machine Learning, NoSQL, Object-Oriented Software Development, Software Engineering, SQL Databases, Data Streaming, Azure Data Factory, Apache Spark, Containerization, Kubernetes, AWS Glue, Integration Frameworks, Apache Kafka, Apache Nifi, Data Management, Machine Learning Operations, Data Pipelines, Apache Beam, Docker - **Published:** July 18, 2026 - **Apply:** https://www.careerjet.com/job/us6e06f6cfc0696ceb2877253183ab3c04/eaa ## About the Role * At least two (2) years' experience working in a Data Engineering, Data Science, Software Development or other relevant role. * Professional experience with programming in either Python or an object-oriented programming language. * Strong knowledge of relational and NoSQL-based databases, with significant proficiency in SQL. * Understanding of ETL processes and data modeling concepts. * Exposure to data processing frameworks and tools (examples are but not all required as Apache Spark, Kafka, dlt, dbt), and cloud data services (AWS Glue, Azure Data Factory, GCP Dataflow). Experience with one or more of these tools or services is a plus but not required. * Knowledge of data warehousing, lakehouse architectures, and data modeling concepts. Experience with ML tools such as pytorch is a plus. * Understanding of machine-learning workflows and ability to build feature stores for AI models. * Exposure to containerization (Docker), Kubernetes and continuous integration/continuous deployment (CI/CD). Experience is a plus but not required. * General understanding of AI/ML concepts with the ability and willingness to learn more. * Ability to collaborate with team leadership and Data Engineering, Infrastructure, AI Engineering, Security and Business peers. * Strong problem-solving, communication and teamwork skills. ## Description Mid-level engineer who designs and maintains scalable data pipelines, ETL processes and data platforms to support AI/ML workloads, integrating vector stores and ensuring data quality and compliance., * Implement and maintain scalable batch and streaming data pipelines to ingest, transform and serve data for AI/ML workloads; work with senior engineers and architects on designing pipelines and processes. * Develop ETL/ELT processes using Python and SQL to prepare training, test, and production datasets and feature stores. Experience with big data technologies (Spark, Hadoop) and flow tools (Kafka, NiFi) is a plus but not required. * Build and maintain data warehouses and lakes; integrate with vector stores to support retrieval-augmented generation (RAG) systems. aPartner with more senior engineers to collaborate with AI Data Engineering, IT Data Engineering, Infrastructure, AI Engineering, Security and Business Leaders to deliver features for model training and inference * Implement data validation and quality checks with validation from more senior engineers; maintain documentation of data flows and schemas. * Work with more senior engineers to ensure pipelines meet data quality, observability, security and regulatory compliance standards. * Work with Model Context Protocol (MCP) to integrate into data pipelines and make modifications to existing MCP connections with guidance from more senior engineers. ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [NoSQL Data Modeling for Front-end Developers](https://www.wearedevelopers.com/videos/297-nosql-data-modeling-for-front-end-developers) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [What Industries Outside of AI Are Hiring The Most AI Experts?](https://www.wearedevelopers.com/magazine/98-what-industries-outside-of-ai-are-hiring-the-most-ai-experts) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Transforming Software Development: The Role of AI and Developer Tools](https://www.wearedevelopers.com/magazine/527-transforming-software-development-the-role-of-ai-and-developer-tools)