> Markdown version of [/jobs/ext/3576518-staff-data-platform-engineer](https://www.wearedevelopers.com/jobs/ext/3576518-staff-data-platform-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Data Platform Engineer - **Company:** Kayak - **Location:** Berlin, Germany - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Airflow, Amazon Web Services, Data Analysis, Apache HTTP Server, Cloud Computing, Code Review, Continuous Integration, Data Validation, Information Engineering, Data Infrastructure, Data Security, Data Sharing, Github, Java Virtual Machine (JVM), Python (Programming Language), Machine Learning, Metadata, Data Streaming, Workflow Management Systems, Parquet, Kubernetes, Deployment Automation, Production Code - **Published:** October 4, 2026 - **Apply:** https://www.adzuna.de/details/5910261690 ## About the Role * 7+ years of professional experience in data engineering, with meaningful time spent at a senior or staff level with domain-wide technical scope. * Experience designing and operating lakehouse architectures at scale - including open table formats (e.g., Apache Iceberg), columnar storage (Parquet) and cloud object storage * Experience building and operating streaming data pipelines - including event-driven ingestion, exactly-once delivery semantics, consumer lag management, checkpoint and recovery strategies, and failure handling in production environments. * Hands-on experience with data contracts, schema governance, metadata, or semantic-layer systems. * Strong Python skills and a track record of writing maintainable, testable production code. * Experience deploying and operating data workloads on Kubernetes - including managing containerized infrastructure, resource tuning and health checks. * Proven ability to influence multiple teams, communicate architectural trade-offs, and drive adoption. * Experience mentoring engineers and raising technical standards through reviews, documentation, and reusable patterns. * Comfort taking ownership of broad, ambiguous problem spaces. To stand out: * Distributed query engines such as Trino. * Workflow orchestration tools such as Apache Airflow. * Experience with AWS or an equivalent public cloud provider. * CI/CD and deployment automation (e.g., GitHub Actions). * Working knowledge of Java or another JVM-based language, given the JVM-based nature of several frameworks in this domain. ## Description We are looking for a Staff Data Engineer to join our Data Platform team. Our team builds and operates the shared foundation that powers analytics, machine learning, business intelligence and AI-driven experiences across the entire company. As a senior individual contributor, you will shape how we move, store, govern, and serve data at scale, enabling every downstream team to build faster and with greater confidence. If you enjoy turning ambiguity into clear technical direction and durable solutions, we'd love to hear from you. This role will be required to work from our Berlin office 3 days per week. In this role, you will: * Design and evolve the architecture of KAYAK's shared Data Platform, including near-real-time streaming, lakehouse storage, schema management, semantic layer, and distributed query infrastructure. Make thoughtful trade-offs between latency, correctness, cost, and long-term maintainability. * Deliver high-impact platform initiatives end-to-end - from problem framing and architecture design through implementation, rollout, and operational handoff. * Define and promote technical standards for data contracts, schema evolution, ingestion patterns, and production readiness across platform and domain teams. * Lead high-impact platform initiatives from problem framing and architecture design through implementation, rollout, and operational handoff. * Develop reusable patterns and reference architectures for streaming ingestion, compaction, retention, schema governance, observability, and other recurring data engineering challenges. * Establish reliable observability across the platform, including pipeline monitoring, consumer lag tracking, data quality checks, and alerting. * Collaborate closely with Operations, Security, Engineering, Data Engineering, and Product to evolve the platform, build cross-functional support, and ensure the platform meets the needs of its users. * Drive the semantic layer and metadata strategy that supports consistent and trusted self-service analytics and AI-driven data access. * Evaluate technologies and approaches across streaming, storage, query, orchestration, and cloud infrastructure, balancing scalability, operational complexity, cost, and maintainability. * Coach and mentor engineers through design reviews, code reviews, pairing, and reusable technical guidance. * Own the most complex architectural and operational challenges on the platform, including failure recovery, schema drift, partition management, and performance degradation. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Parquet, Delta, Iceberg & Ducklake - An introduction for developers](https://www.wearedevelopers.com/videos/100075-parquet-delta-iceberg-ducklake-an-introduction-for-developers) - [Fully Orchestrating Databricks from Airflow](https://www.wearedevelopers.com/videos/336-fully-orchestrating-databricks-from-airflow) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Tracking vehicles at scale](https://www.wearedevelopers.com/videos/1999-tracking-vehicles-at-scale) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [The Biggest German Tech Companies](https://www.wearedevelopers.com/magazine/424-the-biggest-german-tech-companies) - [7 Most Popular Web Developer Jobs in Europe](https://www.wearedevelopers.com/magazine/163-7-most-popular-web-developer-jobs-in-europe) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)