Staff Software Engineer, Data Governance & Foundations

Instacart
United States
about 1 month ago
Apply on us.experteer.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Apache HTTP Server Data Governance Data Infrastructure Distributed Systems Software Engineering Large Language Models Apache Spark Data Strategy Data Lakes Apache Flink Apache Kafka
+2 more
Data Management Vertica

Job description

Experteer Overview In this role you will shape Instacart’s data backbone by designing a scalable open lakehouse foundation with governance and multi-engine compute. You will partner with Data Science, ML, Ads, and Security teams to balance reliability, performance, and cost as we scale. You’ll lead architecture decisions and mentor engineers to drive cross-org alignment. This is a chance to advance a modern data platform that powers critical real-time and analytics workloads. Join us to help serve millions of shoppers and unlock AI-enabled data capabilities. Compensation / Benefits * Translate data strategy into a multi-year architecture roadmap and align with leadership * Own the open lakehouse foundation including unified table formats and governance * Define and manage a multi-engine compute portfolio (interactive, batch, streaming) * Drive real-time and streaming infrastructure for high-impact use cases (Ads, Fraud, ML) * Lead architecture reviews, mentor engineers, influence hiring, and communicate trade-offs to stakeholders * Pioneer AI-native data infrastructure by applying LLM/AI to platform lifecycle and automation Tasks * 10+ years of software engineering experience building and operating data infrastructure or distributed systems at production scale * Hands-on expertise with open table formats (e.g., Apache Iceberg, Delta Lake, Hudi) and distributed query/compute engines (e.g., Trino, Spark, ClickHouse) * Experience with event-driven and streaming infrastructure (e.g., Kafka, Flink) for real-time pipelines * Proven ownership of major platform transitions or migrations delivered to production * Ability to build cost/benefit and TCO models and communicate architecture strategy across teams Key requirements * remote-friendly” * flexible work locations * new hire equity grant * annual equity refresh grants * competitive compensation * benefits package

Requirements

hiring, and communicate trade-offs to stakeholders * Pioneer AI-native data infrastructure by applying LLM/AI to platform lifecycle and automation Tasks * 10+ years of software engineering experience building and operating data infrastructure or distributed systems at production scale * Hands-on expertise with open table formats (e.g., Apache Iceberg, Delta Lake, Hudi) and distributed query/compute engines (e.g., Trino, Spark, ClickHouse) * Experience with event-driven and streaming infrastructure (e.g., Kafka, Flink) for real-time pipelines * Proven ownership of major platform transitions or migrations delivered to production * Ability to build cost/benefit and TCO models and communicate architecture strategy across teams Key requirements * remote-friendly” * flexible work locations * new hire equity grant * annual equity refresh grants * competitive compensation * benefits package

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:24 min

The governance failures of centralized data lakes

Mario Meir-Huber · LIVE

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

2:04 min

Comparing offline data analytics with online stream processing

Artem Volk Artem Volk +1 · World Congress 2024

6:24 min

Distributed data lakes and containerized computing clusters

Ulrich Wurstbauer +1 · LIVE

1:59 min

Evolving roles in AI driven software teams

Ignacio Riesgo Ignacio Riesgo +1 · World Congress 2024

Videos

See all

Related articles

See all