> Markdown version of [/jobs/ext/1912524-senior-data-generalist](https://www.wearedevelopers.com/jobs/ext/1912524-senior-data-generalist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Generalist - **Company:** AMO SERVICES - **Location:** Paris, France - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Information Engineering, Data Mining, Data Systems, Distributed Data Store, Apache Hive, Python (Programming Language), Machine Learning, Query Optimization, SQL Databases, Parquet, Data Processing, Delivery Pipeline, Data Lakes, Pyspark, Apache Flink, Spark Streaming, Stream Processing, Data Pipelines - **Published:** August 4, 2026 - **Apply:** https://fr.indeed.com/viewjob?jk=ebbd78331eb30c56 ## About the Role Senior-level experience building and operating data pipelines. * Strong Python and SQL. * Production experience with Python or Scala. * Experience with PySpark, Spark SQL, Spark Streaming, Parquet, and Iceberg. * Strong understanding of batch and stream processing. * Query optimization experience. * Familiarity with data engine internals and distributed data systems. * Experience with data lakes, table/storage formats, and derived datasets., * Experience with Flink or similar stream processing systems. * Experience with ML training or inference pipelines. * Familiarity with ranking, recommendations, embeddings, or entity resolution. * Experience with privacy-sensitive data systems, GDPR workflows, or data deletion/export pipelines. * Experience reading or making scoped changes in Rust-backed systems. ## Description Build and maintain batch and streaming data pipelines. * Move production data into archival/data lake storage safely and observably. * Improve quality, reliability, and observability of derived datasets. * Design validations for freshness, completeness, schema correctness, and data quality. * Build reusable datasets that reduce one-off data extraction and aggregation work. * Optimize queries and data processing jobs for correctness, latency, and cost. * Prototype in the right tools, then integrate scoped production experiments with amo's systems where needed. ## Related Videos - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) - [Parquet, Delta, Iceberg & Ducklake - An introduction for developers](https://www.wearedevelopers.com/videos/100075-parquet-delta-iceberg-ducklake-an-introduction-for-developers) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Tracking vehicles at scale](https://www.wearedevelopers.com/videos/1999-tracking-vehicles-at-scale) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)