Principal Data Engineer - AI

Anaplan

United States

Hybrid

USD 170,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anaplan is seeking a Principal Data Engineer to own the data architecture and drive scalable data pipelines across our AI-enabled planning platform. You will lead end-to-end data systems from ingestion to governance, enabling real-time and batch analytics across massive datasets.

You will collaborate with analytics, product, and platform teams to model customer metrics and build robust data infrastructure using modern tools. Hybrid work with onsite days for some US-based employees.

Qualifications

  • Extensive data engineering experience in complex environments.
  • Deep understanding of vector, NoSQL, and document stores for AI infrastructure.
  • Hands-on building and shipping large-scale data platforms in production.
  • Strong experience with distributed data processing frameworks (Spark/Flink/Hadoop).
  • Proficient with message brokers and streaming (Kafka/Kinesis).
  • End-to-end data pipeline lifecycle including Airflow/Dagster workflows.
  • Experience with cloud data warehouses (Snowflake/BigQuery/Redshift) and data lakes (Databricks/Delta Lake/Iceberg).
  • Advanced SQL and Python; solid software development practices (testing, CI/CD, IaC).

Responsibilities

  • Lead architecture, design, and deployment of scalable data systems into prod.
  • Develop end-to-end ETL/ELT pipelines and data ingestion frameworks.
  • Build storage and processing layers for analytics workloads (data lakes/warehouses).
  • Create feature pipelines for enterprise datasets with batch and streaming balance.
  • Optimize distributed queries and data transformations for low latency.
  • Implement data quality, governance, and observability across assets.
  • Collaborate with analytics, product, and platform teams on data models.
  • Evaluate new tools and frameworks within the modern data stack.

Skills

Data engineering
Data architecture
Big data platforms
Spark/Flink/Hadoop
Kafka/Kinesis
Airflow/Dagster
Cloud data warehouses
Snowflake
BigQuery
Redshift
SQL
Python
CI/CD
IaC

Tools

Databricks
Delta Lake
Iceberg
Snowflake
BigQuery
Redshift
Airflow
Dagster

Job description

At Anaplan, we are a team of innovators focused on optimizing business decision‑making through our leading AI‑infused scenario planning and analysis platform so our customers can outpace their competition and the market.

What unites Anaplanners across teams and geographies is our collective commitment to our customers’ success and to our Winning Culture.

Our customers rank among the who’s who in the Fortune 50. Coca-Cola, LinkedIn, Adobe, LVMH and Bayer are just a few of the 2,400+ global companies who rely on our best‑in‑class platform.

Our Winning Culture is the engine that drives our teams of innovators. We champion diversity of thought and ideas, we behave like leaders regardless of title, we are committed to achieving ambitious goals, and we love celebrating our wins – big and small.

Supported by operating principles of being strategy‑led, values‑based and disciplined in execution, you’ll be inspired, connected, developed and rewarded here. Everything that makes you unique is welcome; join us and let’s build what’s next – together!

We’re seeking a Principal Data Engineer who can work across the full stack of Anaplan’s data platform, setting the technical direction for how we ingest, transform, store, serve, and govern data at scale. You will build highly performant, robust data pipelines that process massive volumes of data in real‑time and batch. This foundational work empowers business users to leverage vast datasets in their planning workflows and forms the bedrock for our advanced analytics and AI initiatives. You’ll need deep knowledge of distributed computing, data architecture, and strong software engineering skills to tackle complex, high‑scale data challenges.

This role is open to candidates located in the Eastern or Central time zones. Employees who live within commuting distance of one of our offices will be expected to work onsite two days per week as part of our hybrid work model

Your Impact
  • Lead the data architecture, design, and deployment of scalable, high‑throughput Big Data systems into production environments.
  • Architect, deploy, and manage the foundational data systems that underlie modern AI infrastructure, including vector, NoSQL, and document databases.
  • Develop end‑to‑end data engineering solutions, including robust ETL/ELT pipelines, API services, and data ingestion frameworks.
  • Design and build the storage and processing layers powering our analytics workloads: data lakes, data warehouses, distributed file systems, and real‑time streaming architectures.
  • Engineer feature‑rich context pipelines that process large‑scale enterprise data, balancing batch and streaming patterns seamlessly.
  • Optimize and scale large distributed queries and data transformations to ensure high performance and low latency for end users.
  • Implement data quality frameworks to measure and ensure data integrity, reliability, and governance across all data assets.
  • Collaborate with analytics, product, and platform teams to build data models that capture the semantics of customer metrics, hierarchies, and relationships.
  • Stay current with the modern data stack and big data landscape, evaluating new tools, distributed computing frameworks, and database technologies for potential adoption.
Your Skills
  • Extensive data engineering experience, demonstrating a strong track record of hands‑on execution and delivery in complex data environments.
  • Deep practical understanding of the database ecosystems that power AI and machine learning infrastructure (e.g., Vector databases, NoSQL, and Document stores).
  • Hands‑on experience building, scaling, and shipping large‑scale data platforms in production.
  • Deep practical experience with distributed data processing frameworks (e.g., Apache Spark, Flink, Hadoop).
  • Strong expertise in message brokers and event streaming platforms (e.g., Apache Kafka, Kinesis).
  • End‑to‑end exposure to data pipeline lifecycle development, including extensive experience with workflow orchestration tools (e.g., Apache Airflow, Dagster).
  • Hands‑on expertise with cloud data warehouses (e.g., Snowflake, BigQuery, Redshift) and data lake architectures (e.g., Databricks, Delta Lake, Apache Iceberg).
  • Advanced SQL skills and proficiency in Python.
  • Strong background in modern software development practices (testing, code review, CI/CD, Infrastructure as Code).
Desirable
  • Extensive, progressive experience leading technical projects and mentoring engineering teams.
  • Hands‑on experience with cloud‑native infrastructure (AWS, GCP, or Azure).
  • Experience implementing data observability, monitoring, and alerting frameworks at scale.
  • Familiarity with Anaplan or similar enterprise planning platforms.

#LI-SP1

Our Commitment to Diversity, Equity, Inclusionand Belonging (DEIB)

We believe attracting and retaining the best talent and fostering an inclusive culture strengthens our business. DEIB improves our workforce, enhances trust with our partners and customers, and drives business success. Build your career in a place where diversity, equity, inclusion and belonging aren’t just words on paper – this is what drives our innovation, it’s how we connect, and it contributes to what makes us a market leader. We believe in a hiring and working environment where all people are respected and valued, regardless of gender identity or expression, sexual orientation, religion, ethnicity, age, neurodiversity, disability status, citizenship, or any other aspect which makes people unique. We hire you for who you are, and we want you to bring your authentic self to work every day!

We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform essential job functions, and receive equitable benefits and all privileges of employment. Please contact us to request accommodation.

Candidate data processed during our recruitment activities is handled in accordance with our Candidate Privacy Notice. This may include the use of artificial intelligence or automated tools to assist our team in evaluating qualifications.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Data Engineer - AI
Principal Data Engineer - AI

Anaplan • Philadelphia

Hybrid
USD 180,000 - 240,000
Principal Data Scientist - AI
Principal Data Scientist - AI

Anaplan • Philadelphia

Hybrid
USD 180,000 - 240,000
Senior Solution Consultant - AI Specialist (Pre-Sales)
Senior Solution Consultant - AI Specialist (Pre-Sales)

Anaplan Inc • Minneapolis (MN)

On-site
USD 112,000 - 160,000
Senior Solution Consultant - AI Specialist (Pre-Sales)
Senior Solution Consultant - AI Specialist (Pre-Sales)

Anaplan Inc • San Francisco (CA)

Hybrid
USD 139,000 - 188,000
Senior Solution Consultant - AI Specialist (Pre-Sales)
Senior Solution Consultant - AI Specialist (Pre-Sales)

Anaplan Inc • New York (NY)

On-site
USD 139,000 - 188,000
Senior Solution Consultant - AI Specialist (Pre-Sales)
Senior Solution Consultant - AI Specialist (Pre-Sales)

Anaplan • New York (NY)

On-site
USD 139,000 - 188,000
Senior Solution Consultant - AI Specialist (Pre-Sales)
Senior Solution Consultant - AI Specialist (Pre-Sales)

Anaplan • San Francisco (CA)

On-site
USD 139,000 - 188,000
Principal Solution Architect, Professional Services - Finance
Principal Solution Architect, Professional Services - Finance

Anaplan Inc • Minneapolis (MN), Northern (KY)

Hybrid
USD 137,000 - 197,000
Principal Solution Architect, Professional Services - Finance
Principal Solution Architect, Professional Services - Finance

Anaplan • San Francisco (CA)

On-site
USD 171,000 - 232,000
Senior Solution Consultant - AI Specialist (Pre-Sales)
Senior Solution Consultant - AI Specialist (Pre-Sales)

Anaplan • Minneapolis (MN)

On-site
USD 112,000 - 160,000