> Markdown version of [/jobs/ext/2930406-staff-software-engineer-data-products](https://www.wearedevelopers.com/jobs/ext/2930406-staff-software-engineer-data-products). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Software Engineer, Data Products - **Company:** Dune - **Location:** London, UK (Remote available) - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Big Data, Data Systems, Software Debugging, Distributed Systems, Python (Programming Language), Standard Sql, SQL Databases, Rust (Programming Language), Parquet, Backend, Kotlin, Data Lakes, Apache Flink, Apache Kafka, Spark Streaming, Stream Processing - **Published:** September 16, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=bdf1f49ae2e28659 ## About the Role * You are a backend engineer who has gone deep on data systems, or a data engineer who became a strong software engineer. You ship production services, not only pipelines * You have built or materially extended orchestration and scheduling systems, and can explain precisely what breaks at scale and why * You have handled schema evolution and data correctness in a system with real consumers downstream, where a breaking change has a cost * You have built or operated stateful stream processing in production (Flink,Kafka Streams, Spark Structured Streaming, RisingWave, Materialize, Feldera) * You have strong SQL and modeling skills on large datasets, and an interest in how the query engine underneath actually executes your work * You have solid computer science fundamentals and distributed systems understanding * You debug independently and drive root cause analysis to a fix that holds * You use AI tools well enough that they have changed how you work, you understand their failure modes and dislike ai-slop. * You communicate clearly in writing and get the best out of a distributed team, * Deep experience with a transformation framework such as dbt or SQLMesh: specifically, having hit its limits and built beyond them * Data lake formats such as Parquet, Iceberg or Deltalog * Stateful stream processing in production (Flink, Kafka Streams, Spark Structured Streaming) * Experience at a company where the data is the product ## Description Data Products builds and owns datasets end to end: from raw chain data through decoding to the 3000+ models and 4 petabytes we curate, share directly with customers, and replicate into their warehouses. The role will focus on the lifecycle of building high quality data: orchestrating thousands of interdependent models, propagating schema changes without breaking downstream consumers, propagating corrections. That is a software architecture problem in a data domain. This role is a hybrid: a backend engineer who thinks in systems and contracts, working on data. You will be the engineer we hand ambiguous product requirements to, and will come back with a design, a sequence, and work the team can pick up, while building the hardest parts yourself., * Design and build the control plane for our curated data lifecycle: dependency-aware orchestration, backfills, restatements, retries, partial failure, and recovery * Decide, dataset by dataset, whether the answer is a model, a service or a job, and own that architecture through production * Design the contracts between ingestion and curation so a dataset can be reasoned about end to end * Build alerting and data quality signals that catch real problems and stay quiet otherwise, so on-call is about incidents rather than noise * Work across Go, Kotlin, Rust, Python and SQL, choosing the right tool rather than the familiar one * Break large problems into work other engineers can own, and sequence it so we ship something useful early ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Parquet, Delta, Iceberg & Ducklake - An introduction for developers](https://www.wearedevelopers.com/videos/100075-parquet-delta-iceberg-ducklake-an-introduction-for-developers) - [Kotlin Multiplatform - True power of native code reuse](https://www.wearedevelopers.com/videos/4-kotlin-multiplatform-true-power-of-native-code-reuse) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)