> Markdown version of [/jobs/ext/361899-data-engineer](https://www.wearedevelopers.com/jobs/ext/361899-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Manus Technology Group - **Location:** Eindhoven, Netherlands - **Experience:** Expert - **Salary:** €5,000.0 - €6,200.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Confluence, JIRA, Microsoft Azure, Cloud Engineering, Information Engineering, Electronic Data Capture, Python (Programming Language), Machine Learning, SQL Databases, Technical Data Management Systems, Google Cloud, Backend, Git, Containerization, Kubernetes, Information Technology, Terraform, Data Pipelines, Docker - **Published:** June 20, 2026 - **Apply:** https://nl.indeed.com/viewjob?jk=a266af258f6c6870 ## About the Role Do you have experience in Terraform?, · Strong background in Computer Science, Data Engineering, or a related field. · 5+ years of experience in a data-focused or backend role. · Advanced proficiency in Python and SQL. · Experience handling multimodal data: such as video, audio, or time-series sensor data. · Deep expertise in architecting cloud-based data storage and processing solutions. · Excellent communication skills: you can bridge the gap between technical data requirements and high-level AI goals. * Excellent proficiency in both spoken and written English. A Plus Would Be · Experience with Robotics (ROS/ROS2) or AI training pipelines. · Knowledge of local data capture systems and edge-to-cloud synchronization. · Familiarity with containerization (Docker, Kubernetes) and Terraform. · Experience with Git, Jira, and Confluence. ## Description As a Data Engineer at MANUS, you will work together with a bright, multidisciplinary team of engineers to build a robust data foundation for the future of robotics and AI. Our goal is to facilitate the collection of high-quality, multimodal datasets to train embodied AI: AI that interacts with and understands the physical world. In this role, you are the subject matter expert on the team regarding data orchestration and infrastructure. You will work closely with our Embodied AI Engineers to determine what data needs to be collected: ranging from video and audio to various sensor streams: and how it should be structured to be "AI-ready." Your focus will be on the end-to-end data lifecycle: from orchestrating local collection at the source to managing the efficient transition and storage of that data in the cloud. You have full responsibility for your own projects and are expected to work independently. Your Responsibilities · Multimodal Data Orchestration: Acting as the team's authority on capturing diverse data types (video, sensors, etc.) and ensuring they are perfectly structured for machine learning training loops. · Pipeline Management: Designing and maintaining the pipelines that manage local data collection and the subsequent upload and ingestion into our cloud ecosystem. · Cloud Architecture: Building and maintaining scalable cloud environments (AWS, Google Cloud, or Azure) optimized for storing and processing massive, multi-modal datasets. · Synchronization & Alignment: Solving the complex challenge of time-syncing video frames with high-frequency sensor data to ensure "frame-perfect" training sets. · Collaboration: Working side-by-side with AI engineers to align data collection strategies with model requirements. ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Collaboration Quantified: Lessons from Open Source Developer Networks](https://www.wearedevelopers.com/videos/1422-collaboration-quantified-lessons-from-open-source-developer-networks) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) - [Industrializing your Data Science capabilities](https://www.wearedevelopers.com/videos/178-industrializing-your-data-science-capabilities) ## Related Articles - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best Companies in the Netherlands: Top 25 Companies in 2023 ](https://www.wearedevelopers.com/magazine/193-best-companies-in-the-netherlands-top-25-companies-in-2023) - [Software Developer Salary in The Netherlands [2023]](https://www.wearedevelopers.com/magazine/217-software-developer-salary-in-the-netherlands-2023) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)