Senior Data Engineer - ML Platform

Sovendus GmbH
Frankfurt am Main, Germany
2 months ago
Apply on indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Part-time (≤ 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Languages
English
Job source

Tech stack

Airflow Amazon Web Services Amazon S3 Data Analysis Apache HTTP Server Microsoft Azure Cloud Computing Code Review Information Engineering Amazon DynamoDB Apache Hive Python (Programming Language)
+15 more
Machine Learning Redis Azure Machine Learning SQL Databases Sql Optimization Apache Spark Technical Debt Data Layers Data Lakes Pyspark Real Time Data Apache Kafka Machine Learning Operations Terraform Stream Processing

Job description

  • Work in a modern, cloud-only AWS stack - EMR, EKS, S3, Athena/Glue, SageMaker, DynamoDB - orchestrated with Airflow and managed as code with Terraform, with a clear path to bringing every remaining piece fully onto this stack.
  • Shape the Kafka backbone that carries user-journey events from upstream services to the data lake, ML training, and analytics consumers.
  • Move data at scale with Spark SQL on Delta Lake and Apache Iceberg - ACID storage, time-travel, and concurrent-write reliability on top of S3.
  • Build the data layer for our recommendation engine - Athena schemas and views for model training, plus the real-time data assets (Redis, DynamoDB) that the recommendation service reads at request time.
  • Push AI-assisted development and operations further - we invest heavily in AI tooling for development, code review, documentation and operations, and we’re actively deepening that. You’ll help shape how the team works.
  • Steward the platform’s long-term health - drive version upgrades, reduce tech debt, and keep the stack on the leading edge.

Requirements

Do you have experience in Terraform?, * 5+ years of experience in data engineering or ML platform work in production environments

  • You write strong Python code, focused on production-grade applications (not just notebooks or scripts)
  • Advanced SQL (Athena / Spark SQL; SQL-first data transformations)
  • Kafka experience (producers, consumers, stream processing)
  • Solid Cloud experience (AWS preferred, or GCP/Azure)
  • Fluent English for daily collaboration within an international team.
  • Based in Spain with valid work authorization

It´s a plus if you have :

  • Spark & data lake ecosystem - Spark/PySpark (job structuring & tuning), plus lakehouse formats like Delta Lake or Iceberg
  • Orchestration & MLOps - Airflow, SageMaker or similar tools
  • ML data workflows - experience with feature pipelines, training datasets, and online serving

About the company

Become part of one of Europe’s leading networks for checkout and digital marketing services. More than 3,000 European partner shops trust our high-quality e-commerce solutions to enhance their business. How? By innovatively integrating artificial intelligence into online marketing. Now it’s your turn: Enrich our teams with your ideas, energy, and personality., 1 Team. 18 Nationalities. 50% Women & 50% Men. Hundreds of Opportunities!

Our story began in 2008 in Germany: Oliver Stoll founded the company “Gutschein-Connection”. This marked the start of successful growth that has made us the leading network for vouchers and special offers. In 2011, the company was renamed Sovendus, and today, Sovendus has 145 employees who have made this tremendous growth possible. This is also due to our great diversity: our teams consist of roughly equal numbers of women and men, with a total of 18 nationalities and an average age of 35. This includes not only permanent employees but also working students and interns.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

Videos

See all

Related articles

See all