AI Data Engineer

FLEXJET, LLC
Cleveland, OH, United States
about 1 month ago
Apply on us.experteer.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Cloud Computing Information Systems Computer Programming Data Architecture Information Engineering Data Governance Extract Transform Load (ETL) Data Warehousing Python (Programming Language)
+13 more
Standard Sql Unstructured Data Data Processing Authentication Authorization Accounting (Server Program) Data Storage Management Google Cloud Apache Spark Pandas Information Technology Non-relational Database Machine Learning Operations Data Pipelines Automation Anywhere

Job description

Experteer Overview In this role you will design and maintain data infrastructure to power AI models, collaborating with data scientists, ML engineers, and software teams to deliver reliable, scalable data pipelines. You will ensure data quality, governance, and performance across systems for AI workflows. The position offers impact through enabling model training, validation, and inference at scale. This is an opportunity to shape data architecture in a cross-functional, AI-driven environment. Compensation / Benefits * Design, build, and maintain data pipelines for AI and ML workflows * Collect, clean, and preprocess structured and unstructured data * Develop and manage datasets for model training, validation, and inference * Collaborate with ML engineers and data scientists to support model development * Ensure data quality, integrity, and availability across systems * Optimize data storage and retrieval for performance and scalability * Implement data governance, security, and compliance best practices * Monitor and troubleshoot data pipeline issues Tasks * Bachelors degree in Computer Science, Data Engineering, Information Systems, or related field (or equivalent experience) * Strong programming skills in Python and/or SQL * Understanding of ETL/ELT, data modeling, data warehousing * Familiarity with machine learning workflows and data requirements * Experience with data processing tools (e.g., Pandas, Spark) * Knowledge of relational and non-relational databases * Basic understanding of cloud platforms (AWS, Azure, or Google Cloud) Key requirements *

Requirements

Monitor best practices * Monitor and troubleshoot data pipeline issues Tasks * Bachelors degree in Computer Science, Data Engineering, Information Systems, or related field (or equivalent experience) * Strong programming skills in Python and/or SQL * Understanding of ETL/ELT, data modeling, data warehousing * Familiarity with machine learning workflows and data requirements * Experience with data processing tools (e.g., Pandas, Spark) * Knowledge of relational and non-relational databases * Basic understanding of cloud platforms (AWS, Azure, or Google Cloud) Key requirements *

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

6:10 min

Unlocking free learning credits via Google Cloud Innovators

Asrar Asrar · World Congress 2024

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:58 min

Analyzing production code coverage data using pandas

Markus Harrer Markus Harrer · World Congress 2021

1:43 min

AWS infrastructure stack and data flow pipeline overview

Artem Volk Artem Volk +1 · World Congress 2024

Videos

See all

Related articles

See all