Graduate Data Engineer

SRG
Marlow, UK
1 day ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Internship / Graduate position
Employment type
Full-time (> 32 hours)
Experience level
Starter
Experience required
2 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Big Data Data Architecture Dataspaces Data Systems Data Visualization DevOps Python (Programming Language) Machine Learning NumPy
+23 more
Power BI Cloud Services DataOps Software Engineering TypeScript Google Cloud Feature Engineering Large Language Models Multi-Agent Systems Prompt Engineering Apache Spark Generative AI Pandas Matplotlib Pyspark Information Technology Data Analytics Apache Kafka Non-relational Database Data Management Machine Learning Operations Tools for Reporting Data Pipelines

Job description

As a Graduate Data Engineer, you will build and maintain scalable data pipelines in Palantir Foundry for advanced reporting and analytics while collaborating with cross-functional teams as part of the BTS Data & Analytics team. You will work closely with key stakeholders in Engineering, Product, GTM, and other groups to help build scalable data solutions that support key metrics, reporting, and insights. You will assist in ensuring teams have access to reliable, accurate data as our company grows. You will have the opportunity to support projects that enable self-serve insights, helping teams make data-driven decisions, while learning from experienced team members and developing your technical and business skills., * Build and maintain data pipelines, leveraging PySpark and/or Typescript within Foundry, to transform raw data into reliable, usable datasets. Familiarity with Palantir Foundry, PySpark, Kafka, TypeScript, PowerBI preferable.

  • Assist in preparing and optimizing data pipelines to support machine learning and AI model development, ensuring datasets are clean, well-structured, and readily usable by Data Science teams.
  • Support the integration and management of feature engineering processes and model outputs into Foundry’s data ecosystem, helping enable scalable deployment and monitoring of AI/ML solutions as you develop your skills in this area.
  • Engaged in gathering and translating stakeholder requirements for key data models and reporting, with a focus on Palantir Foundry workflows and tools.
  • Participate in developing and refining dashboards and reports in Foundry to visualize key metrics and insights as you grow your data visualization skills.
  • Collaborate with Product, Engineering, and GTM teams to align data architecture and solutions, learning to support scalable, self-serve analytics across the organization.
  • Have some prompt engineering experience with large language models, including writing and evaluating complex multi-step prompts.
  • Continuously develop your understanding of the company’s data landscape, including Palantir Foundry’s ontology-driven approach and best practices for data management.

Requirements

  • You have a degree in Computer Science, Engineering, Mathematics, or similar, or have similar work experience.
  • Having up to 2 years of experience building data pipelines at work or through internships is helpful.
  • You can write clear and reliable Python/PySpark code.
  • You are familiar with popular analytics tools (like pandas, numpy, matplotlib), big data frameworks (like Spark), and cloud services (like Palantir, AWS, Azure, or Google Cloud).
  • You have a deep understanding of data models, relational and non-relational databases, and how they are used to organize, store, and retrieve data efficiently for analytics and machine learning.
  • Knowing about software engineering methods, including DevOps, DataOps, or MLOps, is also a plus., * Master’s degree in engineering (such as AI/ML, Data Systems, Computer Science, Mathematics, Biotechnology, Physics), or minimum 2 years of relevant technology experience.
  • Experience with Generative AI (GenAI) and agentic systems will be considered a strong plus.
  • Have a proactive and adaptable mindset: willing to take initiative, learn new skills, and contribute to different aspects of a project as needed to drive solutions from start to finish, even beyond the formal job description.
  • Show a strong ability to thrive in situations of ambiguity, taking initiative to create clarity for yourself and the team, and proactively driving progress even when details are uncertain or evolving.
  • Hybrid working policy: Currently, our client expects all staff to be in their Marlow-based office at least 3 days a week from Jan 2026.

About the company

SRG are working with a leading pharmaceutical company based in Marlow. Our client develops and manufacture an impressive portfolio of aesthetics brands and products. Our client is committed to driving innovation and providing high-quality products and services.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

1:25 min

Replacing NumPy with cuPy for straightforward GPU acceleration

Paul Graham Paul Graham · World Congress 2025

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

Videos

See all

Related articles

See all