> Markdown version of [/jobs/ext/1164031-data-engineer](https://www.wearedevelopers.com/jobs/ext/1164031-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** RIMES Technologies Corporation - **Location:** Paris, France - **Experience:** Starter - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Data Analysis, Big Data, Databases, Data Validation, Information Engineering, Extract Transform Load (ETL), Data Systems, Distributed Computing Environment, Python (Programming Language), Operational Data Store, Performance Tuning, SQL Databases, Data Streaming, Workflow Management Systems, Data Ingestion, Apache Spark, Caching, Pyspark, Data Analytics, Data Management, Data Pipelines, Databricks - **Published:** July 3, 2026 - **Apply:** https://fr.indeed.com/viewjob?jk=70b97454c734d639 ## About the Role Note: Experience with Palantir Foundry or Databricks is a strong plus but not required. If you bring solid data engineering fundamentals in Python/PySpark, SQL, and modern ELT patterns, we'll support a fast ramp-up on both platforms., * 1-3 years in data engineering or analytics engineering with end-to-end pipeline delivery in production. * Proficiency in Python & PySpark for distributed data processing. * Strong SQL for analytical and transformation logic. * Data modeling skills for both analytics and operational use cases. * Experience with data ingestion from APIs, databases, external feeds, and real-time sources. * Solid grasp of data quality, testing, observability, lineage, and governance practices. * Comfort working with large datasets and distributed compute using modern ELT patterns. Nice To Have: * Databricks or cloud-native compute with compute pushdown. * Palantir Foundry: pipelines/transforms, Code Repos, Ontology, and operational applications. * Spark execution concepts: partitions, shuffles, caching, and performance optimization. * Experience with financial or enterprise operational data. * Experience with AI-assisted ETL/ELT or data quality tooling. * Familiarity with streaming frameworks and/or orchestration tools. ## Description We're looking for a Data Engineer to own data onboarding and build scalable, reliable data pipelines that power analytics, operational workflows, and data-driven decisions across Rimes. You'll work closely with data producers, analysts, and product teams to ingest, transform, and operationalize data across Palantir Foundry and Databricks as our two target data platforms., * Ingest & onboard datasets from internal systems, APIs, databases, files, external providers, and real-time feeds. * Build and operate scalable ETL/ELT pipelines using Python, PySpark, SQL, and Foundry pipeline tooling; schedule and automate batch/stream refreshes. * Model and operationalize data (e.g., defining entities/relationships) to support analytics and operational applications in collaboration with domain experts. * Ensure trust in data through testing, data quality checks, observability/alerting, lineage, and compliant access controls. * Collaborate with analysts and product teams to translate business requirements into robust data solutions and clear data contracts/SLOs. What Success Looks Like (First 3-6 Months): * You onboard and productionize new data sources with reliable refresh (scheduled or real-time). * You deliver trusted, well-documented datasets consumed by analytics and operational teams. * Key business entities are clearly modeled and discoverable. * Pipelines have meaningful monitoring and alerting, with reduced failures/re-runs. * You contribute to standards/templates that speed up future onboarding. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [HTTP headers that make your website go faster](https://www.wearedevelopers.com/videos/1676-http-headers-that-make-your-website-go-faster) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story)