Staff Software Engineer, Data Products
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+5 more
Job description
Data Products builds and owns datasets end to end: from raw chain data through decoding to the 3000+ models and 4 petabytes we curate, share directly with customers, and replicate into their warehouses.
The role will focus on the lifecycle of building high quality data: orchestrating thousands of interdependent models, propagating schema changes without breaking downstream consumers, propagating corrections. That is a software architecture problem in a data domain. This role is a hybrid: a backend engineer who thinks in systems and contracts, working on data.
You will be the engineer we hand ambiguous product requirements to, and will come back with a design, a sequence, and work the team can pick up, while building the hardest parts yourself., * Design and build the control plane for our curated data lifecycle: dependency-aware orchestration, backfills, restatements, retries, partial failure, and recovery
- Decide, dataset by dataset, whether the answer is a model, a service or a job, and own that architecture through production
- Design the contracts between ingestion and curation so a dataset can be reasoned about end to end
- Build alerting and data quality signals that catch real problems and stay quiet otherwise, so on-call is about incidents rather than noise
- Work across Go, Kotlin, Rust, Python and SQL, choosing the right tool rather than the familiar one
- Break large problems into work other engineers can own, and sequence it so we ship something useful early
Requirements
- You are a backend engineer who has gone deep on data systems, or a data engineer who became a strong software engineer. You ship production services, not only pipelines
- You have built or materially extended orchestration and scheduling systems, and can explain precisely what breaks at scale and why
- You have handled schema evolution and data correctness in a system with real consumers downstream, where a breaking change has a cost
- You have built or operated stateful stream processing in production (Flink,Kafka Streams, Spark Structured Streaming, RisingWave, Materialize, Feldera)
- You have strong SQL and modeling skills on large datasets, and an interest in how the query engine underneath actually executes your work
- You have solid computer science fundamentals and distributed systems understanding
- You debug independently and drive root cause analysis to a fix that holds
- You use AI tools well enough that they have changed how you work, you understand their failure modes and dislike ai-slop.
- You communicate clearly in writing and get the best out of a distributed team, * Deep experience with a transformation framework such as dbt or SQLMesh: specifically, having hit its limits and built beyond them
- Data lake formats such as Parquet, Iceberg or Deltalog
- Stateful stream processing in production (Flink, Kafka Streams, Spark Structured Streaming)
- Experience at a company where the data is the product
Benefits & conditions
- A competitive salary and equity package . Both salary and equity is top 25% of companies in the space
- Our employee equity scheme has world-class employee-friendly terms with a heavily discounted strike price (~90%) and a 10-year exercise window
- 5 weeks PTO + local public holidays (that can be swapped to suit you)
- A fully remote-first approach within a distributed team with flexible working hours; you structure your own day
- Say goodbye to meeting overload! We believe in a healthy mix of async and sync work, so you can focus on what truly matters-no more wasted time on endless meetings!
- Good health is important, so we offer private medical insurance, dental & vision as standard
- We believe in paid parental leave to help you celebrate this important milestone, transition to your new life, and bond with your new baby. We offer 16 weeks to primary and 6 weeks to secondary caregivers, fully paid. Plus a 2-week part-time phased return at full pay to help you get used to your new (and slightly more complex!) schedule
- Quarterly offsites in various exciting locations as a company or team to connect, work together and have fun (so far in Tuscany Berlin Austria and Athens ).
- On top of this each person gets a yearly travel allowance to connect and co-work with someone or a team of people for a few days.
- An allowance for your at-home setup, to ensure you are happy, comfortable and productive. If you prefer a local co-working space, we’ll pay for your desk.
- Work with some of the best people you’ll ever get to meet!
- And of course, you get some awesome Dune swag!
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
Dev Digest 120 - Apple and peers
Highest Paying Tech Companies for Developers
Navigating the AI Shift