> Markdown version of [/jobs/ext/325265-forward-deployed-data-engineer](https://www.wearedevelopers.com/jobs/ext/325265-forward-deployed-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Forward Deployed Data Engineer - **Company:** The Ellison Institute Of Technology (eit) - **Location:** Oxford, UK - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Cloud Computing, Cloud Storage, Data Validation, Information Engineering, Data Infrastructure, Data Loss, Linux System Administration, Message Broker, Message Queuing Telemetry Transport (MQTT), Rapid Prototyping Process, Cloud Services, Robotic Automation Software, Data Streaming, Management of Software Versions, Data Layers, Process Control Systems, Apache Kafka, Machine Learning Operations, Stream Processing, Data Pipelines - **Published:** June 3, 2026 - **Apply:** https://www.careerjet.co.uk/jobad/gba60fceb8bbf0e23aef2dd07cffbdc4d1 ## About the Role * You have strong programming experience in Python, and value code quality, reliability, and readability as much as performance. * You have experience working on cloud compute platforms, containers and Linux environments. * You think in systems and own them end-to-end - from device output to APIs - and embrace long-term engineering rather than one-off scripts. * Hands-on experience building data pipelines for physical systems: autonomous vehicles, robotics, clinical/lab instruments, industrial control systems, or similar environments. * Experience with real-time data streaming: Kafka or equivalent message brokers, MQTT or similar protocols, and the challenges that come with high-frequency, low-latency data capture. Great to Also Have * Familiarity with live video or high-bandwidth media streams as data engineering problems is a strong plus. * Experience in automated lab, clinical lab, or life sciences environments - understanding of instrument APIs, lab protocols, and the data quality expectations of scientific workflows. * Comfort with time-series data from sensors and control systems, including sampling rates, data loss handling, and operational (driving live systems) vs analytical (modelling) use cases of the same data. * Understanding of closed-loop control systems and the data infrastructure needed to support real-time decision making. ## Description Our platform connects physical hardware - robotic systems, ambient sensor rigs, lab hardware - to a cloud-native data platform that captures, stores and services high-frequency multimodal data. The platform spans an edge-to-cloud telemetry stack (MQTT, Kafka, multimedia streams), REST and streaming APIs, and a growing set of hardware integrations across multiple sites. As a Data Engineer, you'll work directly alongside hardware engineers, scientists, and the core platform team to build the data layer that makes our autonomous operations possible. This is a hands-on, high-impact role for someone who is comfortable combining hardware data streams with reproducible, version-controlled, and well-structured data pipelines, and is comfortable bringing their own expertise into diverse groups. Day-to-Day, You Might * Integrate new lab instruments and robotic systems into the platform - writing device APIs, defining event schemas, and validating data at the edge before it reaches cloud storage. * Build and maintain high-throughput, low-latency ingestion pipelines for streaming data: live video feeds, sensor telemetry, robot joint state, and instrument control signals. * Work across the edge-to-cloud stack - from configuring edge devices and MQTT brokers through to Kafka topics, cloud storage, and the APIs that expose data to scientists and downstream ML pipelines. * Design and evolve the common data model for hardware execution and scientific outcome data, with a strong focus on schema stability, versioning, and provenance. * Collaborate with scientists and hardware engineers to turn raw instrument output into research-ready, schema-validated data for model training. * Contribute to an engineering culture that values maintainability, testing, robust system design, and deep collaboration, but allows flexibility for rapid prototyping and responsiveness to changing landscapes., You'll be one of a nimble team building infrastructure that directly enables autonomous science at scale. The hardware is real, the data volumes are high, the use cases span from live lab control to training foundation models - and the decisions you make about schemas, protocols, and pipeline architecture now will shape how the platform grows across new sites and hardware categories in 2027 and beyond. ## Related Videos - [How to Avoid LLM Pitfalls - Mete Atamel and Guillaume Laforge](https://www.wearedevelopers.com/videos/1328-how-to-avoid-llm-pitfalls-mete-atamel-and-guillaume-laforge) - [Why Your AI Agent Keeps Hallucinating Your Data: Building Deterministic Context Layers](https://www.wearedevelopers.com/videos/2055-why-your-ai-agent-keeps-hallucinating-your-data-building-deterministic-context-layers) - [How to Benchmark Your Apache Kafka](https://www.wearedevelopers.com/videos/76-how-to-benchmark-your-apache-kafka) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Fault Tolerance and Consistency at Scale: Harnessing the Power of Distributed SQL Databases](https://www.wearedevelopers.com/videos/1146-fault-tolerance-and-consistency-at-scale-harnessing-the-power-of-distributed-sql-databases) - [Python-Based Data Streaming Pipelines Within Minutes](https://www.wearedevelopers.com/videos/1233-python-based-data-streaming-pipelines-within-minutes) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)