Data Engineer

Lupa Pets
London, UK
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Airflow Amazon Web Services Information Engineering Data Infrastructure Data Transformation Python (Programming Language) Delivery Pipeline Apache Spark Build Management Data Lakes Machine Learning Operations Data Pipelines
+1 more
Docker

Job description

  • Design and build data pipelines from raw ingestion through to clean, modelled, production-ready datasets that engineering and product teams rely on
  • Own data quality across your domain, defining standards, instrumenting checks, and resolving failures with urgency
  • Monitor pipeline health in production and respond to issues proactively, not reactively
  • Collaborate with engineering and product to understand data requirements and turn them into robust, scalable solutions
  • Help define and evolve the architecture of our data platform as we scale across new geographies and product surfaces
  • Document your work clearly so the team can understand, extend, and maintain what you build

Requirements

Do you have experience in Spark?, * Multiple years of hands-on data engineering experience in a professional setting

  • Strong Python skills and a demonstrated ability to build pipelines from scratch, not just extend existing ones
  • Solid knowledge of data modelling and practical experience with Apache Spark, AWS services, Docker, and Airflow
  • Experience in a startup or fast-growth environment where you’ve made pragmatic decisions under uncertainty
  • Scala experience (nice to have, we’re happy to invest in teaching it if your foundations are strong)
  • Familiarity with dbt, Delta Lake, or similar modern data transformation tooling (nice to have)
  • Exposure to ML pipelines or feature stores (nice to have)

As a person, you:

  • Have a proactive, ownership-oriented mindset. You don’t wait to be told something is broken
  • Take genuine pride in craft and correctness, and hold yourself to a high bar
  • Collaborate well with engineering teams who consume your data and communicate clearly across functions
  • Thrive in environments where you’re expected to build from scratch, not inherit a tidy queue of tickets
  • Are excited to work in-person from our Paddington, London HQ (or equivalent location)

What does success look like in 6 months?

  • You’ve built and shipped at least two meaningful pipelines end-to-end
  • You own a defined area of the data platform and are making architectural calls with confidence
  • The team trusts your judgement without needing to review every decision

About the company

Lupa is building a category defining product the industry has never seen before.

We’re the AI-native operating system for veterinary practices and pet parents, replacing the fragmented, clunky systems vets have tolerated for years with a single intelligent platform for scheduling, client communication, clinical documentation, and AI-driven care guidance. Practices run more efficiently, vets get back to doing what they love, and pet parents feel more connected to their animals’ health than ever.

The traction speaks for itself: founded in 2023 and already one of Europe’s top 100 AI startups, with a team of 50 people, 10x growth in twelve months, the UK market leader, and now charging hard into the US and Europe. We have $25M in funding, 1M+ pets on the platform, and a buzzing HQ in Paddington, London. We’ve attracted exceptional people from the likes of Palantir, Google, DeepMind, BCG, Meta, and AWS and we’re just getting started.

This is a rare chance to join a rocket ship at exactly the right moment and we’re looking for exceptional people to help us fly it.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · WWC Europe 2026

3:24 min

The governance failures of centralized data lakes

Mario Meir-Huber · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

6:24 min

Distributed data lakes and containerized computing clusters

Ulrich Wurstbauer +1 · LIVE

Videos

See all

Related articles

See all