Principal Data Engineer - AI

Anaplan
United States
5 days ago
Apply on job-boards.greenhouse.io
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours

Tech stack

Artificial Intelligence Airflow Amazon Web Services Apache HTTP Server Microsoft Azure Big Data BigQuery Cloud Database Code Review Databases Continuous Integration Data Architecture
+36 more
Information Engineering Data Governance Data Infrastructure Data Integrity Extract Transform Load (ETL) Data Transformation Data Systems Data Warehousing Distributed File Systems Digital Assets Distributed Computing Environment Distributed Systems Apache Hadoop Python (Programming Language) Message Broker NoSQL DataOps Anaplan Software Engineering Data Streaming Web Services Software Organization Cloud Platform System Data Ingestion Sql Optimization Snowflake Apache Spark Build Management Data Lakes Infrastructure Automation Frameworks Low Latency Apache Flink Apache Kafka Data Pipelines Amazon Redshift Databricks

Job description

We’re seeking a Principal Data Engineer who can work across the full stack of Anaplan’s data platform, setting the technical direction for how we ingest, transform, store, serve, and govern data at scale. You will build highly performant, robust data pipelines that process massive volumes of data in real-time and batch. This foundational work empowers business users to leverage vast datasets in their planning workflows and forms the bedrock for our advanced analytics and AI initiatives. You’ll need deep knowledge of distributed computing, data architecture, and strong software engineering skills to tackle complex, high-scale data challenges. This role is open to candidates located in the Eastern or Central time zones. Employees who live within commuting distance of one of our offices will be expected to work onsite two days per week as part of our hybrid work model Your Impact

  • Lead the data architecture, design, and deployment of scalable, high-throughput Big Data systems into production environments.
  • Architect, deploy, and manage the foundational data systems that underlie modern AI infrastructure, including vector, NoSQL, and document databases.
  • Develop end-to-end data engineering solutions, including robust ETL/ELT pipelines, API services, and data ingestion frameworks.
  • Design and build the storage and processing layers powering our analytics workloads: data lakes, data warehouses, distributed file systems, and real-time streaming architectures.
  • Engineer feature-rich context pipelines that process large-scale enterprise data, balancing batch and streaming patterns seamlessly.
  • Optimize and scale large distributed queries and data transformations to ensure high performance and low latency for end users.
  • Implement data quality frameworks to measure and ensure data integrity, reliability, and governance across all data assets.
  • Collaborate with analytics, product, and platform teams to build data models that capture the semantics of customer metrics, hierarchies, and relationships.
  • Stay current with the modern data stack and big data landscape, evaluating new tools, distributed computing frameworks, and database technologies for potential adoption., We believe attracting and retaining the best talent and fostering an inclusive culture strengthens our business. DEIB improves our workforce, enhances trust with our partners and customers, and drives business success. Build your career in a place where diversity, equity, inclusion and belonging aren’t just words on paper - this is what drives our innovation, it’s how we connect, and it contributes to what makes us a market leader. We believe in a hiring and working environment where all people are respected and valued, regardless of gender identity or expression, sexual orientation, religion, ethnicity, age, neurodiversity, disability status, citizenship, or any other aspect which makes people unique. We hire you for who you are, and we want you to bring your authentic self to work every day!

We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform essential job functions, and receive equitable benefits and all privileges of employment. Please contact us to request accommodation.

Fraud Recruitment Disclaimer

It has come to our attention that fraudulent and fictitious job opportunities are being circulated on the Internet. Prospective candidates are being contacted by certain individuals, mainly through telephone calls, emails and correspondence, claiming they are representatives of Anaplan. The main purpose of these correspondences and announcements is to obtain privileged information from individuals.

Anaplan does not:

  • Extend offers to candidates without an extensive interview process with a member of our recruitment team and a hiring manager via video or in person.
  • Send job offers via email. All offers are first extended verbally by a member of our internal recruitment team whenever possible and then followed up via written communication.

Requirements

  • Extensive data engineering experience, demonstrating a strong track record of hands-on execution and delivery in complex data environments.
  • Deep practical understanding of the database ecosystems that power AI and machine learning infrastructure (e.g., Vector databases, NoSQL, and Document stores).
  • Hands-on experience building, scaling, and shipping large-scale data platforms in production.
  • Deep practical experience with distributed data processing frameworks (e.g., Apache Spark, Flink, Hadoop).
  • Strong expertise in message brokers and event streaming platforms (e.g., Apache Kafka, Kinesis).
  • End-to-end exposure to data pipeline lifecycle development, including extensive experience with workflow orchestration tools (e.g., Apache Airflow, Dagster).
  • Hands-on expertise with cloud data warehouses (e.g., Snowflake, BigQuery, Redshift) and data lake architectures (e.g., Databricks, Delta Lake, Apache Iceberg).
  • Advanced SQL skills and proficiency in Python.
  • Strong background in modern software development practices (testing, code review, CI/CD, Infrastructure as Code).

Desirable

  • Extensive, progressive experience leading technical projects and mentoring engineering teams.
  • Hands-on experience with cloud-native infrastructure (AWS, GCP, or Azure).
  • Experience implementing data observability, monitoring, and alerting frameworks at scale.
  • Familiarity with Anaplan or similar enterprise planning platforms.

Benefits & conditions

As set forth in Anaplan’s Equal Employment Opportunity policy, we do not discriminate on the basis of any protected group status under any applicable law. Gender Select… Are you Hispanic/Latino? Select… Race & Ethnicity Definitions

If you believe you belong to any of the categories of protected veterans listed below, please indicate by making the appropriate selection. As a government contractor subject to the Vietnam Era Veterans Readjustment Assistance Act (VEVRAA), we request this information in order to measure the effectiveness of the outreach and positive recruitment efforts we undertake pursuant to VEVRAA. Classification of protected categories is as follows:

A “disabled veteran” is one of the following: a veteran of the U.S. military, ground, naval or air service who is entitled to compensation (or who but for the receipt of military retired pay would be entitled to compensation) under laws administered by the Secretary of Veterans Affairs; or a person who was discharged or released from active duty because of a service-connected disability.

A “recently separated veteran” means any veteran during the three-year period beginning on the date of such veteran’s discharge or release from active duty in the U.S. military, ground, naval, or air service.

An “active duty wartime or campaign badge veteran” means a veteran who served on active duty in the U.S. military, ground, naval or air service during a war, or in a campaign or expedition for which a campaign badge has been authorized under the laws administered by the Department of Defense.

An “Armed forces service medal veteran” means a veteran who, while serving on active duty in the U.S. military, ground, naval or air service, participated in a United States military operation for which an Armed Forces service medal was awarded pursuant to Executive Order 12985. Veteran Status Select…

About the company

At Anaplan, we are a team of innovators focused on optimizing business decision-making through our leading AI-infused scenario planning and analysis platform so our customers can outpace their competition and the market.

What unites Anaplanners across teams and geographies is our collective commitment to our customers’ success and to our Winning Culture.

Our customers rank among the who’s who in the Fortune 50. Coca-Cola, LinkedIn, Adobe, LVMH and Bayer are just a few of the 2,400+ global companies who rely on our best-in-class platform.

Our Winning Culture is the engine that drives our teams of innovators. We champion diversity of thought and ideas, we behave like leaders regardless of title, we are committed to achieving ambitious goals, and we love celebrating our wins - big and small.

Supported by operating principles of being strategy-led, values-based and disciplined in execution, you’ll be inspired, connected, developed and rewarded here. Everything that makes you unique is welcome; join us and let’s build what’s next - together!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on job-boards.greenhouse.io
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:00 min

Separating dataset creation from low-level software implementation steps

Jan Zawadzki · World Congress 2022

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

1:31 min

Essential AI and human skills for future teams

Alexander Weißhaupt Alexander Weißhaupt +1 · World Congress 2025

Videos

See all

Related articles

See all