Senior Python Data Engineer (AI & Data Migration) Job ID: JP054841

Itproposal
Aalter, Belgium
25 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Languages
Dutch, English, French

Tech stack

Artificial Intelligence Apache HTTP Server Information Engineering Extract Transform Load (ETL) Data Transformation Data Migration Data Transformation Services Python (Programming Language) Performance Tuning SAS (Software) SAS/Base Software Engineering
+11 more
SQL Databases Workflow Management Systems Working Model 2D Parquet Macros GitHub Copilot Multi-Agent Systems Apache Spark Pandas Pyspark Data Pipelines

Job description

Senior Python Data Engineer (AI-Assisted Data Migration)

We are looking for a Senior Python Data Engineer to support the migration of legacy SAS-based ETL processes to a modern Python data platform using Parquet, PyArrow, Pandas, Polars, and Apache Spark. You will play a key role in developing AI-assisted migration solutions, optimizing agent-driven workflows, and ensuring high-quality, scalable data transformation processes.

Key Responsibilities

  • Design, develop, and validate Python-based ETL and data transformation solutions.
  • Review and optimize AI-generated Python code for performance, quality, and maintainability.
  • Build and enhance data pipelines using Pandas, Polars, PyArrow, Parquet, and PySpark.
  • Improve ETL/ELT processes, data quality, schema management, and performance.
  • Contribute to AI-assisted software development and multi-agent orchestration workflows.
  • Enhance workflow automation, validation processes, and human-in-the-loop quality assurance.
  • Collaborate with architects, data engineers, and SAS experts to support complex migration scenarios.
  • Ensure robust testing, documentation, and operational reliability of migration solutions.

Required Skills \& Experience

  • 5 years of experience in Python Data Engineering.
  • Strong expertise in Python, Pandas, Polars, PyArrow, Parquet, Apache Spark/PySpark, and Python performance tuning.
  • Experience with ETL/ELT, data engineering, data quality, schema management, and testing frameworks.
  • Experience using AI-assisted development tools (e.g., GitHub Copilot or similar).
  • Knowledge of agentic workflows, multi-agent systems, workflow orchestration, and AI-driven automation.
  • Familiarity with SAS Base (DATA Step, PROC SQL, Macros, ETL patterns) is a strong advantage.

Languages

  • Active knowledge of Dutch , French , and English.

Location

  • Brussels, Belgium
  • Hybrid working model.

Requirements

Apache, SQL, Python

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on be.engineering.jobs

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:03 min

Exploring declarative and procedural macro subtypes in Rust environments

Mykhailo Maidan · LIVE

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · WWC 2024

3:33 min

Refactoring data science workflows using Rapids QDF and Pandas

Paul Graham Paul Graham · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all