Sr. Data Engineer

BambooHR LLC

Utah

Hybrid

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
401k with company match
Generous paid time off
Hybrid work schedule

Job summary

BambooHR is seeking a Senior Data Engineer to design, build, and operate scalable data platforms supporting AI-powered HR solutions. You will help create knowledge layers, RAG readiness, and enterprise data products.

You will collaborate across data analysts, ML engineers, and stakeholders to deliver robust pipelines, models, and semantic architectures with a focus on governance, security, and scalable cloud infrastructure.

Qualifications

  • Hands-on experience building AI agents and agentic workflows.
  • Expert-level Python development for scalable data, AI, and knowledge engineering solutions.
  • Advanced SQL development and query optimization across transactional, analytical, and semantic data stores.
  • Experience with Databricks, Spark, and distributed data processing frameworks.
  • Experience designing canonical data models, semantic models, ontologies, taxonomies, and reusable domain-oriented data products.
  • Strong understanding of metadata management, data cataloging, lineage, governance, and data discovery practices.

Responsibilities

  • Collaborate with data analysts, data scientists, ML engineers, business stakeholders, and AI engineers to enable trusted use of enterprise data and knowledge assets.
  • Design, develop, and maintain scalable data pipelines using Python, SQL, PySpark, and modern data engineering frameworks.
  • Build and optimize data lake, lakehouse, warehouse, data mart, and semantic data architectures.
  • Design, build, and maintain an enterprise knowledge layer that unifies structured and unstructured information for AI and analytics workloads.
  • Develop and maintain canonical data models, facts, dimensions, feature datasets, business entities, metadata models, and domain-specific data products.
  • Design pipelines that ingest documents, knowledge bases, APIs, SaaS applications, event streams, and other enterprise content into analytics and AI-ready formats.

Skills

Python development
SQL development
PySpark
Databricks
Spark
Data pipelines
Lakehouse
Data warehouse
Semantic layer
Ontology & taxonomy
Metadata management
Terraform
CI/CD pipelines
Git workflows
Monitoring & observability
Knowledge graphs
RAG / embeddings
AWS
Vector databases

Education

Bachelor's degree in Computer Science or related quantitative field

Tools

Databricks
Spark
Terraform
AWS
Git

Job description

Please Note: This is a Utah-based hybrid position which will require some regular in-office days each week. Additionally, employment with BambooHR is contingent on passing both a background and credit check.

AI at BambooHR

At BambooHR, we’re all about setting people free to do great work, and we believe AI is a powerful partner in that mission. We’re leaning into intelligent tools to streamline our workflows, giving us more time for high-impact innovation. We look for curious, forward-thinking people who are ready to explore how AI can elevate their work and help us reimagine the future of HR.

Essential Job Duties

As a Senior Data Engineer, you will play a key role in designing, building, and operating scalable data platforms, analytics systems, AI/ML infrastructure, and the enterprise knowledge layer that powers intelligent applications and AI agents.

You'll help extract, load, and transform structured and unstructured enterprise data into trusted, searchable, and reusable knowledge assets that enable retrieval-augmented generation (RAG), knowledge graphs, semantic search, AI agents, and advanced analytics. We’ll rely on your expertise across data, AI, and knowledge engineering to develop reliable systems that make organizational knowledge accessible at scale.

Your ability to leverage AI to build performant data platforms, agentic workflows, and enterprise knowledge systems will be critical to your success.

You will:

  • Collaborate with data analysts, data scientists, ML engineers, business stakeholders, and AI engineers to enable trusted use of enterprise data and knowledge assets.
  • Design, develop, and maintain scalable data pipelines using Python, SQL, PySpark, and modern data engineering frameworks.
  • Build and optimize data lake, lakehouse, warehouse, data mart, and semantic data architectures.
  • Design, build, and maintain an enterprise knowledge layer that unifies structured and unstructured information for AI and analytics workloads.
  • Develop and maintain canonical data models, facts, dimensions, feature datasets, business entities, metadata models, and domain-specific data products.
  • Design pipelines that ingest documents, knowledge bases, APIs, SaaS applications, event streams, and other enterprise content into analytics and AI-ready formats.
  • Build pipelines for extracting, chunking, enriching, classifying, and embedding unstructured content.
  • Design and manage vector databases and embedding pipelines to support semantic search and Retrieval-Augmented Generation (RAG).
  • Build and optimize retrieval pipelines including hybrid search, metadata filtering, reranking, and context assembly.
  • Design and implement Knowledge Graph and Graph RAG architectures to model relationships between enterprise entities, documents, people, products, customers, and business processes.
  • Develop entity extraction, relationship extraction, ontology, taxonomy, and metadata enrichment pipelines to improve knowledge discovery.
  • Translate business requirements into scalable data models, semantic models, knowledge schemas, ERDs, data flow diagrams, and analytics and AI-ready architectures.
  • Design and manage cloud-based data and AI infrastructure (Databricks preferred), including development, staging, and production environments.
  • Design evaluation frameworks for retrieval quality, grounding accuracy, hallucination reduction, answer relevance, and AI system performance.
  • Partner with data governance to implement MCP servers, metadata management, data cataloging, lineage, governance, and access controls that improve discoverability and trust of enterprise knowledge.
  • Participate in peer code reviews, pull requests, architecture reviews, and engineering standards.
  • Document data pipelines, knowledge pipelines, AI architectures, semantic models, infrastructure, and operational procedures.
  • Define infrastructure as code and support CI/CD pipelines for data, AI, and knowledge engineering systems.
  • Ensure enterprise data privacy, security governance, and responsible AI practices.
  • Continuously improve platform scalability, resilience, retrieval performance, and operational efficiency.
  • Contribute to the evolution of enterprise data, AI, and knowledge platform architecture and engineering best practices.
What You Need to Get the Job Done

(If you don’t have everything, we still encourage you to apply.)

Collaboration & Business Engagement
  • Ability to translate business problems into scalable data, AI, and knowledge engineering solutions.
  • Experience working cross-functionally with technical and non-technical stakeholders.
  • Ability to quickly learn new business domains and emerging analytics and AI technologies.
  • Strong communication skills with the ability to explain complex technical concepts
Core Technical Skills
  • Hands-on experience building AI agents and agentic workflows.
  • Expert-level Python development for building scalable data, AI, and knowledge engineering solutions.
  • Advanced SQL development and query optimization across transactional, analytical, and semantic data stores.
  • Strong experience with Databricks, Spark, and distributed data processing frameworks.
  • Experience designing and implementing scalable data pipelines for both structured and unstructured enterprise data using Databricks and PySpark.
  • Deep understanding of lakehouse, data warehouse, semantic layer, and enterprise knowledge architecture principles.
  • Experience designing canonical data models, semantic models, ontologies, taxonomies, and reusable domain-oriented data products.
  • Strong understanding of metadata management, data cataloging, lineage, governance, and data discovery practices.
  • Experience developing document ingestion, enrichment, chunking, and indexing pipelines to support AI-powered search and retrieval.
  • Familiarity with embedding generation, vector indexing, and semantic retrieval concepts for Retrieval-Augmented Generation (RAG) systems.
  • Experience with cloud-native data platforms (AWS preferred) and modern storage architectures.
  • Experience implementing Infrastructure as Code (Terraform or similar), CI/CD pipelines, and automated deployment practices.
  • Experience building observable, secure, and resilient data platforms with monitoring, testing, and operational best practices.
  • Proficiency with Git-based development workflows and collaborative software engineering practices

Beyond technical skills, we’re looking for someone who is:

  • A systems thinker who enjoys connecting data, knowledge, and AI.
  • Passionate about building trusted enterprise knowledge that powers intelligent experiences.
  • Curious about emerging AI architectures and rapidly evolving technologies.
  • Analytical and pattern-oriented.
  • Creative in designing scalable data and AI solutions.
  • Detail-oriented and persistent in solving complex engineering challenges.
  • Comfortable working in a fast-paced, collaborative environment.
  • Committed to continuous learning and engineering excellence.
  • Bachelor's degree in Computer Science, Information Systems, Engineering, Mathematics, or a related quantitative field (or equivalent practical experience).
What Will Make Us REALLY Love You
  • Experience designing and evolving enterprise-scale data platform architectures.
  • Experience working with, developing, and deploying MCP servers.
  • Experience implementing event-driven architectures, streaming data pipelines, and Change Data Capture (CDC) technologies.
  • Experience with Infrastructure as Code (Terraform, CloudFormation, or similar) and cloud automation.
  • Experience building and operationalizing machine learning pipelines for training, validation, deployment, monitoring, and observability.
  • Familiarity with enterprise data governance, metadata management, and data stewardship practices and tools.
  • Experience implementing data security, privacy, and regulatory compliance frameworks.
  • Experience supporting real-time analytics, low-latency data processing, or AI inference systems.
  • Familiarity with common business metrics and data models across finance, sales, marketing, product, customer success, and operations.
What You'll Love About Us
  • A Great Company Culture that has been recognized by multiple organizations like Inc, and Salt Lake Tribune
  • Comprehensive health, life, and disability insurance
  • Generous leave policies that include 4 weeks of vacation, 12 company holidays, parental leave, and volunteer time off so you can enjoy quality of life
  • 401k plans with up to 6% company match
  • $2000 Paid-Paid Vacation bonus
  • EAP through Headspace
  • Check out all our benefits that benefit you
About Us

At BambooHR, we're building something different: we're building a people intelligence platform that transforms HR and sets people free to do great work! We're a proven market leader driving innovation while building lasting success through thoughtful, sustainable growth. Here, you'll find a place that champions growth: both professional and personal, both individual and collective.

We invest in potential, giving you the space to stretch your capabilities and turn good ideas into reality while providing the safety net of a supportive, values-driven culture. Our approach combines meaningful work with meaningful lives, offering competitive benefits, professional development, and the flexibility to thrive both in and outside the office.

What sets us apart isn't just what we do, but how we do it: with openness, integrity, and a shared commitment to doing the right thing. Join us in creating HR software that makes work better for everyone, while we make work better for you.

BambooHR is committed to the full inclusion of all qualified individuals and will ensure that persons with disabilities are provided reasonable accommodations throughout the hiring process. If you would like to request accommodations, please let your recruiter know.

BambooHR is An Equal Opportunity Employer--M/F/D/V

Because our team members are trusted to handle sensitive information, we require all candidates that receive and accept employment offers to complete a background check before being hired.

For information on California Privacy Policy, click here.

Our process utilizes AI as an assistant to efficiently process and analyze candidate data. Recruiters and hiring managers maintain full oversight and accountability, ensuring that all final selection and rejection decisions are human-made and based solely on objective job qualifications. Please see our General Privacy Notice and California Privacy Notice for more details.

See our AI Guidelines for Candidates for details on how BambooHR uses AI in recruiting, how we expect candidates to use AI, and what is not allowed.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer — AI-Driven Data Platform Architect
Senior Data Engineer — AI-Driven Data Platform Architect

BambooHR LLC • Utah

On-site
Sr. Data Governance Engineer
Sr. Data Governance Engineer

Bamboohr17 • Utah

Hybrid
USD 100,000 - 130,000
Health, life, and disability insurance
401k plans with up to 6% company match
Generous leave policies including vacation and holidays
Staff AI Platform Engineer - Hybrid
Staff AI Platform Engineer - Hybrid

BambooHR LLC • Utah

On-site
Distinguished Architect, Data Technology
Distinguished Architect, Data Technology

BambooHR • Draper (UT)

On-site
USD 130,000 - 160,000
Comprehensive health, life, and disability insurance
401k plans with up to 6% company match
Generous leave policies including 4 weeks of vacation
+1
Principal Software Architect
Principal Software Architect

BambooHR • Draper (UT)

On-site
USD 130,000 - 170,000
Comprehensive health, life, and disability insurance
Generous leave policies
401k plans with up to 6% company match
+2
Sr.Digital Analyst II
Sr.Digital Analyst II

Bamboohr17 • Arizona

Hybrid
USD 90,000 - 130,000
Health insurance
401(k) with company match
Paid vacation bonus
+2
Sr. Product Manager, Core Data Platform
Sr. Product Manager, Core Data Platform

Bamboohr17 • Utah

On-site
USD 120,000 - 180,000
Health insurance
Life insurance
Disability insurance
+4
Sr. Product Manager, Core Data Platform
Sr. Product Manager, Core Data Platform

Socket.dev • Utah

Hybrid
USD 140,000 - 190,000
Health, life, and disability insurance
401k with company match
Generous paid leave and holidays
+1
Sr. Product Manager, Core Data Platform
Sr. Product Manager, Core Data Platform

BambooHR LLC • Utah

Hybrid
USD 120,000 - 180,000
Health, life, and disability insurance
401k with company match
Vacation 4 weeks + holidays
Chief Architect: AI-Driven SaaS Platform Leader
Chief Architect: AI-Driven SaaS Platform Leader

BambooHR LLC • Utah

On-site