Principal Data Engineer

Vomela

United States

On-site

USD 180,000 - 200,000

Full time

10 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health Care Plan (Medical, Dental &amp
Retirement Plan (401k)
Life Insurance
Paid Time Off
Short Term & Long-Term Disability
Training & Development
Wellness Resources

Job summary

Vomela is seeking a Principal Data Engineer to own the data platform end-to-end, defining data strategy and delivering scalable, trusted data solutions. You will lead architecture decisions, write production code, and mentor others while shaping the organization’s data roadmap.

The role emphasizes Microsoft Fabric as the data platform and requires strong capabilities in streaming, semantic modeling, and data governance to support enterprise reporting and analytics across business units.

Qualifications

  • Microsoft Fabric: Lakehouses, Notebooks, Dataflows Gen2, Event streams, Semantic Models, Direct Lake mode
  • Power BI: report development and semantic model design
  • SQL Server / Azure SQL: query optimization, schema design
  • Azure DevOps: Git-based CI/CD for data pipelines
  • Demonstrated use of AI coding assistants in production engineering workflow
  • Ability to critically evaluate, edit, and improve AI-generated code

Responsibilities

  • Design and implement dimensional models, star schemas, and snowflake schemas with rigor
  • Build and maintain semantic models that serve as the single source of truth for business reporting
  • Implement Slowly Changing Dimension (SCD) strategies appropriate to each domain
  • Own master data engineering: golden record patterns, source-of-record authority, cross-system identity resolution
  • Establish and enforce data modeling standards across the team
  • Design and operate real-time and near-real-time pipelines using streaming technologies (Kafka, Confluent Cloud, Fabric Eventstreams)
  • Relentlessly drive down data staleness in non-streaming scenarios through scheduling, incremental load optimization, and pipeline orchestration design
  • Own performance tuning across the full stack — query optimization, partition strategy, indexing, Delta table compaction, semantic model refresh efficiency, and Direct Lake readiness
  • Apply operational engineering discipline: pipeline observability, alerting, SLA definition, failure recovery, and capacity planning
  • Design and implement controls for sensitive data (financials, PII, HIPAA)
  • Build robust, scalable, observable pipelines — watermark-based incremental loads, CDC patterns, batch and streaming architectures
  • Ensure pipelines are idempotent, recoverable, and production-hardened
  • Serve as the senior technical voice in code review — your approval carries weight
  • Translate business requirements into semantic models and report-layer artifacts for non-technical users
  • Define and enforce data contracts — schema stability, access patterns, SLAs — for each consumer class
  • Own the developer experience of the platform: discoverability, documentation, and onboarding

Skills

Dimensional modeling
Data warehousing
SQL
Performance tuning
Data pipelines
Real-time streaming
Kafka
Fabric
Power BI
Delta Lake

Tools

Microsoft Fabric
Kafka
Confluent Cloud
Fabric Eventstreams
Notebooks
Pipelines
Direct Lake
Power BI
SQL Server/Azure SQL

Job description

At Vomela our greatest asset is our people. As a full-service visual communications company, we are looking for creative and intellectual thinkers that work with our customers to create compelling brand solutions and foster meaningful connections. And while you're focused on creating big things for global and local brands, we will help you build a career you can be passionate about.

Pay Range:$180 - 200k USD

Job Summary

The Principal Data Engineer is the highest-performing contributor on our data engineering team - the person who sets the technical bar, owns the data platform end to end, and delivers work that others study. You'll define and execute data strategy at the engineering level, operating as the technical point of the spear for how the organization builds, scales, and trusts its data. You write production code. You design the architecture. You solve the problems that block everyone else. You mentor without being asked, influence without authority, and deliver without handholding. You're a force multiplier and you're hungry to shape not just the platform, but the broader data strategy of the business.

Microsoft Fabric is our data platform. This role is for someone genuinely energized by the Fabric ecosystem, who tracks its evolution closely and sees its breadth - Lakehouse's, Event streams, Semantic models, Notebooks, Pipelines, Direct Lake as an opportunity, not a constraint. If you're looking for a role where your technical judgment shapes the trajectory of the entire data organization, this is exactly it.

What You'll Do...
  • Design and implement dimensional models, star schemas, and snowflake schemas with rigor
  • Build and maintain semantic models that serve as the single source of truth for business reporting
  • Implement Slowly Changing Dimension (SCD) strategies appropriate to each domain
  • Own master data engineering: golden record patterns, source-of-record authority, cross-system identity resolution
  • Establish and enforce data modeling standards across the team
  • Design and operate real-time and near-real-time pipelines using streaming technologies (Kafka, Confluent Cloud, Fabric Eventstreams) — and know when streaming is the right answer and when it isn't
  • Relentlessly drive down data staleness in non-streaming scenarios through intelligent scheduling, incremental load optimization, and pipeline orchestration design
  • Own performance tuning across the full stack — query optimization, partition strategy, indexing, Delta table compaction, semantic model refresh efficiency, and Direct Lake readiness
  • Apply operational engineering discipline: pipeline observability, alerting, SLA definition, failure recovery, and capacity planning
  • Design and implement controls appropriate for sensitive data (financials, PII, HIPAA, etc.)
  • Build robust, scalable, observable pipelines — watermark-based incremental loads, CDC patterns, batch and streaming architectures
  • Ensure pipelines are idempotent, recoverable, and production-hardened
  • Serve as the senior technical voice in code review — your approval carries weight
Report & Analytics Delivery
  • Translate business requirements into semantic models and report-layer artifacts that non-technical users can trust and navigate
  • Serve as the platform's primary technical interface across consumer groups: Power BI report builders needing trusted, well-modeled semantic layers; AI/ML developers needing governed, feature-ready data surfaces; application developers consuming data via SQL endpoints, REST APIs, or Direct Lake
  • Define and enforce data contracts — schema stability, access patterns, SLAs — for each consumer class
  • Own the developer experience of the platform: discoverability, documentation, and onboarding

At Vomela our greatest asset is our people. As a full-service visual communications company, we are looking for creative and intellectual thinkers that work with our customers to create compelling brand solutions and foster meaningful connections. And while you're focused on creating big things for global and local brands, we will help you build a career you can be passionate about.

Pay Range:$180 - 200k USD

Job Summary

The Principal Data Engineer is the highest-performing contributor on our data engineering team - the person who sets the technical bar, owns the data platform end to end, and delivers work that others study. You'll define and execute data strategy at the engineering level, operating as the technical point of the spear for how the organization builds, scales, and trusts its data. You write production code. You design the architecture. You solve the problems that block everyone else. You mentor without being asked, influence without authority, and deliver without handholding. You're a force multiplier and you're hungry to shape not just the platform, but the broader data strategy of the business.

Microsoft Fabric is our data platform. This role is for someone genuinely energized by the Fabric ecosystem, who tracks its evolution closely and sees its breadth - Lakehouse's, Event streams, Semantic models, Notebooks, Pipelines, Direct Lake as an opportunity, not a constraint. If you're looking for a role where your technical judgment shapes the trajectory of the entire data organization, this is exactly it.

What You'll Do...
  • Design and implement dimensional models, star schemas, and snowflake schemas with rigor
  • Build and maintain semantic models that serve as the single source of truth for business reporting
  • Implement Slowly Changing Dimension (SCD) strategies appropriate to each domain
  • Own master data engineering: golden record patterns, source-of-record authority, cross-system identity resolution
  • Establish and enforce data modeling standards across the team
  • Design and operate real-time and near-real-time pipelines using streaming technologies (Kafka, Confluent Cloud, Fabric Eventstreams) — and know when streaming is the right answer and when it isn't
  • Relentlessly drive down data staleness in non-streaming scenarios through intelligent scheduling, incremental load optimization, and pipeline orchestration design
  • Own performance tuning across the full stack — query optimization, partition strategy, indexing, Delta table compaction, semantic model refresh efficiency, and Direct Lake readiness
  • Apply operational engineering discipline: pipeline observability, alerting, SLA definition, failure recovery, and capacity planning
  • Design and implement controls appropriate for sensitive data (financials, PII, HIPAA, etc.)
ETL / ELT Pipeline Development
  • Build robust, scalable, observable pipelines — watermark-based incremental loads, CDC patterns, batch and streaming architectures
  • Ensure pipelines are idempotent, recoverable, and production-hardened
  • Serve as the senior technical voice in code review — your approval carries weight
Report & Analytics Delivery
  • Translate business requirements into semantic models and report-layer artifacts that non-technical users can trust and navigate
  • Serve as the platform's primary technical interface across consumer groups: Power BI report builders needing trusted, well-modeled semantic layers; AI/ML developers needing governed, feature-ready data surfaces; application developers consuming data via SQL endpoints, REST APIs, or Direct Lake
  • Define and enforce data contracts — schema stability, access patterns, SLAs — for each consumer class
  • Own the developer experience of the platform: discoverability, documentation, and onboarding
Requirements
Required
  • Microsoft Fabric: Lakehouses, Notebooks, Dataflows Gen2, Event streams, Semantic Models, Direct Lake mode
  • Power BI: report development, dataset/semantic model design, DAX proficiency
  • SQL Server / Azure SQL/Postgres: query optimization, schema design, stored procedures
  • Azure DevOps: Git-based development workflows, CI/CD for data pipelines
  • Demonstrated use of AI coding assistants in a production engineering workflow
  • Ability to critically evaluate, edit, and improve AI-generated code and artifacts
  • Clear understanding of where AI accelerates work and where it introduces risk
Preferred Qualifications
  • Familiarity with broader Azure Data Services (Azure Data Factory, Synapse Analytics, ADLS Gen2, Event Hubs) as complementary tooling
  • Experience in a private equity-backed or multi-entity portfolio company environment
  • Exposure to MDM platforms (Profisee, Semarchy, Ataccama, or equivalent)
  • Experience with Confluent Cloud / Apache Kafka for streaming ingestion into Fabric or Synapse
  • Familiarity with cross-tenant Azure / Fabric architecture
  • Background in business analysis, solutions architecture, or pre-sales engineering
  • Microsoft Fabric or Azure Data Engineer certifications
Benefits
  • Health Care Plan (Medical, Dental & Vision)
  • Retirement Plan (401k)
  • Life Insurance (Basic, Voluntary & AD&D)
  • Paid Time Off
  • Short Term & Long-Term Disability
  • Training & Development
  • Wellness Resources
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Data Engineer
Principal Data Engineer

Vomela Company • Northern (KY)

Hybrid
USD 180,000 - 200,000
Health care plan
401k retirement plan
Life insurance
+4
Senior Data Engineer
Senior Data Engineer

Fracht Sweden AB • Houston (TX)

On-site
USD 120,000 - 150,000
Data Engineer
Data Engineer

Sakata Seed America, Inc. • Woodland (CA)

On-site
USD 110,000 - 155,000
Medical, Dental & Vision Insurance
401(k) with Company Match
Paid Vacation & Holidays
Senior Data Engineer
Senior Data Engineer

Fracht Group - North America • Houston (TX)

On-site
USD 120,000 - 150,000
Microsoft Fabric Data Engineer
Microsoft Fabric Data Engineer

St. Petersburg College • Clearwater (FL)

Hybrid
USD 76,000 - 114,000
Microsoft Fabric Data Engineer
Microsoft Fabric Data Engineer

Derextechnologiesinc • Erie

On-site
USD 120,000 - 160,000
Azure Data Engineer
Azure Data Engineer

Vaco Recruiter Services • Dublin (OH)

On-site
USD 110,000 - 140,000
Data Engineer
Data Engineer

Sakataornamentals • Burlington (WA)

On-site
USD 110,000 - 125,000
Health insurance
401(k) Program + Company Match
Holiday Bonus
+2
Data Platform Engineer
Data Platform Engineer

JFC Staffing Companies • Lancaster

On-site
USD 120,000 - 160,000
Microsoft Fabric Sr. Data Engineer
Microsoft Fabric Sr. Data Engineer

2020 Companies, Inc. • Southlake (TX)

On-site
USD 100,000 - 130,000
Competitive salary, paid weekly
Next day pay on demand with DailyPay
Health/Dental/Vision benefits
+6