> Markdown version of [/jobs/ext/747342-young-graduates-databricks-data-ai-engineer](https://www.wearedevelopers.com/jobs/ext/747342-young-graduates-databricks-data-ai-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Young Graduates - Databricks Data & AI Engineer - **Company:** Devoteam - **Location:** Machelen, Belgium - **Contract:** Internship / Graduate position - **Skills:** Clean Code Principles, Java (Programming Language), Application Programming Interfaces (APIs), Artificial Intelligence, Airflow, Amazon Web Services, Amazon S3, Microsoft Azure, Big Data, Cloud Computing, Computer Programming, Databases, Data Governance, Data Infrastructure, Extract Transform Load (ETL), Data Structures, Apache Hadoop, Python (Programming Language), Machine Learning, NoSQL, DataOps, Azure Data Lake, SQL Databases, Retrieval-Augmented Generation, Large Language Models, Data Build Tool (dbt), Apache Spark, Generative AI, Pandas, Data Lakes, Pyspark, Information Technology, Google Bigquery, Apache Kafka, Data Pipelines, Serverless Computing, Docker, Databricks - **Published:** June 19, 2026 - **Apply:** https://www.adzuna.be/details/5676121136 ## About the Role Are you a recent graduate eager to launch your career at the intersection of Big Data and Artificial Intelligence? We are looking for ambitious Young Grads Data & AI Engineers to join our elite Databricks team (50+ FTEs)., * Educational Background: A recent University Degree (Master's preferred) in Computer Science, (Business) Engineering, or a specialized program in Data Science/AI. * Programming & Logic: Strong academic foundation in Python, Scala, or Java. You should be comfortable writing clean, efficient code. * Data Foundations: A solid grasp of SQL and an understanding of how databases work (Relational vs. NoSQL). Familiarity with the concepts of ETL/ELT is a huge plus. * AI Passion: A hunger to build AI things-you follow the latest trends in LLMs and want to know how the data "under the hood" makes them work. * Communication: Excellent interpersonal skills with the ability to collaborate in an agile, multi-disciplinary team. * Languages: Fluency in English and proficiency in Dutch or French is mandatory. Bonus Points * Academic projects or internships involving PySpark, Hadoop, or Kafka. * Experience with Docker or basic Cloud certifications (Azure, AWS, or GCP). * Exposure to dbt (data build tool) or airflow for orchestration. * Knowledge of specialized AI libraries (LangChain, LlamaIndex, or Pandas). ## Description About the Role: Young Graduates - Databricks Data & AI Engineer, * End-to-End Pipeline Engineering: Design, develop, and deploy robust data pipelines using Databricks and Apache Spark. You'll learn to ingest data from diverse sources (APIs, IoT streams, ERPs). * Master the Lakehouse: Get hands-on with Delta Lake to ensure ACID transactions on top of data lakes, ensuring data is reliable, versioned, and "AI-ready." * Hands-on AI Development: Work on implementing Generative AI solutions, including RAG (Retrieval-Augmented Generation) architectures, vector databases, and fine-tuning data for LLMs. * Building for AI: Create the specialized data structures required for Generative AI, including vectorizing data for RAG (Retrieval-Augmented Generation) and building feature stores for Machine Learning models. * Automation & DataOps: Implement Databricks Workflows and CI/CD pipelines to automate data movement, ensuring that high-quality data is always available for business stakeholders and AI agents. * Cloud & Infrastructure: Learn to integrate Databricks with cloud-native services (like Azure Data Lake, AWS S3, or Google BigQuery), mastering the art of cost-effective and scalable compute. * Data Quality & Governance: Use Unity Catalog to implement fine-grained governance, ensuring that the AI solutions we build are secure, compliant, and ethical. ## Related Videos - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Tomorrow's cloud data platforms - fully managed database-as-a-service (DBaaS)](https://www.wearedevelopers.com/videos/254-tomorrow-s-cloud-data-platforms-fully-managed-database-as-a-service-dbaas) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Where to Find Entry-Level Software Engineering Jobs](https://www.wearedevelopers.com/magazine/397-where-to-find-entry-level-software-engineering-jobs) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)