Senior Data Engineer

developrec
Greater London, UK
1 day ago
Apply on www.collegerecruiter.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Information Engineering Data Governance Data Infrastructure Data Integrity Extract Transform Load (ETL) Data Transformation Graph Database Python (Programming Language) Metadata
+11 more
Rule Engine Standard Sql Data Streaming Reinforcement Learning Data Logging Cloud Platform System Large Language Models Data Layers Data Lineage Feature Extraction Data Pipelines

Job description

You will build production-grade data pipelines explicitly aligned to ontologies and semantic models. Your work will ensure that entity definitions, relationships, taxonomies and domain constraints are faithfully represented in data flows, making them reasoning-ready and AI-consumable.

Working within a senior, cross-functional delivery model (consulting, ontology and engineering), you will play a foundational role in building robust semantic layers and enabling high-value AI systems for clients., * Design, build and maintain ETL/ELT pipelines aligned to ontology and knowledge graph structures

  • Implement transformations that respect entity models, relationships, taxonomies and domain constraints
  • Apply semantic enrichment patterns including mapping, harmonisation, linking and feature extraction
  • Deliver high-quality, structured data to downstream AI systems, agents, retrieval layers and decision engines
  • Translate conceptual ontologies into implementable schemas and data flows
  • Partner with ontology architects on entity modelling, semantic definitions, metadata and lineage
  • Deploy pipelines into ontology-aware platforms (e.g. graph databases, semantic layers, Foundry-style systems)
  • Ensure semantic compliance, data integrity and reasoning-readiness

Data Quality, Observability & Lineage

  • Implement robust data quality frameworks (validation, profiling, anomaly detection)
  • Build observability into pipelines (lineage tracking, logging, freshness monitoring, schema drift detection)
  • Ensure alignment with governance, security and industry standards

AI Enablement & Data Serving

  • Build high-quality datasets for retrieval pipelines (RAG), embeddings and conversational agents
  • Create data foundations supporting decision engines, reinforcement learning and value measurement
  • Partner with AI engineers to operationalise pipelines for LLM workflows and agentic systems

Standards, Documentation & Reusability

  • Produce clear documentation for data models, schemas, ontologies and lineage
  • Codify semantic ETL patterns and reusable modelling templates
  • Contribute to internal accelerators, engineering standards and playbooks

Requirements

We are looking for strong data engineering fundamentals combined with demonstrable semantic and ontology experience:

  • 5-8 years’ experience in data engineering, data platform development or data-intensive systems
  • Strong SQL and Python for scalable data transformations and services
  • Experience with at least one major cloud platform (AWS, Azure or GCP)
  • Hands-on experience with semantic or ontology-driven data models, including:
  • RDF/OWL modelling, SHACL validation or ontology tooling
  • Semantic ETL and ontology mapping pipelines
  • Knowledge graph construction, enrichment and query patterns
  • Experience operationalising pipelines for AI systems, LLM workflows or retrieval ecosystems
  • Familiarity with modern data tooling and platform engineering practices
  • Comfortable working in iterative consulting delivery environments with evolving requirements, * High agency - independently drives complex workstreams end-to-end
  • Structured thinker - brings clarity and rigour to ambiguous, messy data domains
  • Collaborative - works effectively with ontology architects, AI engineers and consultants
  • Quality-driven - prioritises correctness, observability, maintainability and semantic integrity
  • Clear communicator - able to explain semantic concepts and data reasoning to non-technical stakeholders
  • Low ego, high ownership - focused on outcomes and value creation

Benefits & conditions

  • You deliver clean, trustworthy, semantically aligned data ready for ontologies and AI layers
  • Ontology architects rely on your pipelines for entity consistency and semantic accuracy
  • AI engineers build faster because your data structures and retrieval layers are reliable and predictable
  • Your semantic ETL patterns and modelling templates are reused across engagements
  • Clients trust your clarity, rigour and dependability in data work underpinning high-value AI systems
  • Your work becomes foundational to the firm’s semantic and agentic engineering capability

Why Join

  • Senior-heavy, engineering-led culture with deep focus on ontologies, knowledge graphs and AI systems
  • Early-stage growth environment backed by significant investment and strong market traction
  • High autonomy, low bureaucracy and meaningful system-building responsibility
  • Opportunity to shape internal standards, accelerators and AI-native products
  • Clear commitment to responsible AI and widening access to advanced technologies
  • Flexible working model with a modern Central London presence
  • Comprehensive health, wellbeing and pension benefits

This is an opportunity to help define how semantic data engineering enables next-generation AI systems, within a firm where clarity, technical depth and real-world outcomes matter.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.collegerecruiter.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:47 min

Comparing Egeria to alternative open metadata solutions

Ferd Scheepers · World Congress 2022

1:58 min

Shifting security permissions from applications to the data layer

Neena Thomas Neena Thomas · World Congress 2026 Europe

9:55 min

Empowering business analysts with a polyglot rule engine

Chris Heilmann +2 · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:08 min

Creating standard APIs via the Egeria open metadata project

Ferd Scheepers · World Congress 2022

1:59 min

Evolving roles in AI driven software teams

Ignacio Riesgo Ignacio Riesgo +1 · World Congress 2024

Videos

See all

Related articles

See all