Data Engineer

briowt

Glendale (CA)

On-site

USD 140,000 - 200,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Medical
Dental
Vision
Life Insurance
401K
Paid Vacation
Paid Holidays

Job summary

Brio Water Technology is seeking a hands-on Data Engineer to design, build, and operate end-to-end data pipelines and a Databricks lakehouse. You will own ingestion, governance, modeling, and ML/AI integrations, collaborating with cross-functional teams to deliver scalable, quality data products.

Responsibilities include building batch and streaming pipelines, developing a governed metrics layer, and deploying models with MLflow while ensuring robust security and auditing.

Qualifications

  • 8+ years in data engineering with production cloud data platform ownership
  • Hands-on with Databricks: Unity Catalog, Delta, Workflows, SQL warehouses
  • Expert dbt and dimensional modeling with a governed metrics layer
  • Strong SQL and Python; proficient with Git, CI/CD, IaC

Responsibilities

  • Ingestion: design/build pipelines from SaaS apps, databases, APIs, and in-house systems
  • Lakehouse and governance: manage a Databricks lakehouse with Unity Catalog
  • Modeling: develop dimensional models and a governed metrics layer
  • Unstructured data & ML: build pipelines turning media into structured data using transcription/OCR/LLM classification
  • AI integration: implement LLM integrations with permissions and audit logging
  • APIs: define requirements and build against application APIs
  • Reporting: create governed dashboards/reports and alerting
  • Observability: monitor pipeline freshness, data quality, and model performance

Skills

Databricks
Unity Catalog
SQL
Python
Data Modeling
CI/CD
Git
Delta Lake

Tools

dbt
Looker
Omni
Sigma
pgvector

Job description

About the company

Home Organizers, Inc.is the parent organization behind a portfolio of well-known home products and services brands, including Closet World, Closets by Design, Brio Water Technology, and others. With decades of experience supporting innovation, design, manufacturing, and customer focused solutions, Home Organizers helps its brands deliver high quality products and services that improve everyday living for households across the country.

Brio Water Technology, a Home Organizers, Inc. company,

is a market leading water products brand that has helped millions stay hydrated through its unique and innovative product line. We offer full home water solutions designed to elevate the way people hydrate, combining sophisticated technology with modern, top tier design to deliver exceptional performance, customer satisfaction, and enhanced functionality.

About the role

We are hiring a strong, hands-on data engineer who works across the full stack, from source ingestion and pipelines through modeling, governance, AI integration, and the reporting layer. You will design, build, and operate the platform yourself. We may add engineers later; the standards you set are the ones they will inherit.

  • Judgment is the job. You weren't hired just to execute. Be judicious. Weigh the cost-benefit against risk on every build-versus-buy decision and every shortcut.
  • Ship with discipline. Every pipeline, model, and metric ships with tests, documentation, version control, and a named owner.
  • Influence over authority. Much of the work depends on partners across the business. You earn their partnership. Run your big calls past me, then move.
What you'll do
  • Ingestion. Design, build, and operate batch and streaming pipelines from SaaS applications, databases, APIs, and custom in-house systems, using managed connectors where they fit and custom code (including change data capture) where they do not.
  • Lakehouse and governance. Build and administer a Databricks lakehouse (medallion architecture, Unity Catalog), including row filters, column masks, tagging, lineage, and identity-provider group sync.
  • Modeling and semantic layer. Develop dimensional models, entity resolution, and a governed metrics layer in dbt. Own testing, documentation, and data quality.
  • Unstructured data and ML. Build pipelines that turn audio, video, documents, images, and sensor or edge-device telemetry into structured data using transcription, OCR, LLM classification, and computer vision. Deploy, version, and monitor models with MLflow; build and maintain vector search indexes.
  • AI integration. Build and maintain LLM integrations (MCP servers, retrieval, text-to-SQL) with both read and write capabilities, enforcing user-level permissions, approval workflows, and audit logging.
  • APIs. Define integration requirements for APIs built by our application engineering team, and build against them.
  • Reporting. Build governed dashboards, reports, scheduled delivery, and alerting on a semantic-model BI platform.
  • Observability. Monitor pipeline freshness, data quality, model performance, and index health, with alerting that surfaces problems early.
Qualifications
  • 8+ years in data engineering, including at least 3 owning a production cloud data platform end to end, from ingestion through the reporting layer.
  • Hands-on production experience with Databricks: Unity Catalog, Delta, Workflows/Lakeflow Jobs, and SQL warehouses.
  • Expert dbt and dimensional modeling: layered projects, tests, documentation, and a semantic or metrics layer (MetricFlow, Metric Views, or LookML).
  • Strong SQL and Python. Comfortable with Git, CI/CD, and infrastructure as code.
  • Built governed BI on a semantic model (Omni, Sigma, Looker, or similar) with row-level security, scheduled delivery, and alerting.
  • Delivered both managed-connector ingestion (Lakeflow Connect, Fivetran, Airbyte) and custom pipelines, including CDC and API sources.
  • Shipped at least one pipeline that turns unstructured data (audio, documents, or images) into structured data using transcription, OCR, LLM classification, or vision models.
  • Operated models in production with MLflow or an equivalent registry, including versioning, scoring jobs, and monitoring.
  • Implemented data access controls: row filters, column masks, PII handling, and group-based permissions synced from an identity provider.
  • Solved entity resolution across multiple source systems (customers, locations, or employees).
  • Built integrations that write back to production systems (in-house applications, CMS, or SaaS APIs) with authorization checks, idempotency, rollback, and audit logging.
  • Consumed APIs built by other teams and written clear requirements for what a data or AI integration needs from them.
  • Able to work directly with senior business leaders to define metrics and explain trade-offs in plain language.
Strongly Preferred
  • Built LLM applications on governed data: MCP servers, text-to-SQL with an evaluation set, and retrieval systems covering embedding models, chunking strategy, hybrid (keyword plus vector) search, metadata and permission filtering, and retrieval evaluation, using Databricks Vector Search, pgvector, or similar.
  • Experience with enterprise search or knowledge platforms (Glean or similar), knowledge graphs, or business glossaries.
  • Data observability tooling (Metaplane, Elementary, Monte Carlo, Lakehouse Monitoring).
  • Change-management automation: preview environments, automated tests, approval workflows, and policy-as-code (OPA, Cedar, or similar).
  • Streaming and IoT data: sensor or edge-device telemetry landed through Structured Streaming, Kafka, or edge-to-cloud pipelines.
  • Multi-cloud work across AWS, Azure, and GCP.
  • Marketing and advertising data (Google Ads, Meta, GA4, call tracking).
  • Home services, franchise, field service, or retail data.
  • Working knowledge of SOC 2 controls and California privacy requirements (CCPA/CPRA).
Benefits / Perks

We believe in recognizing and rewarding our employees for a job well done. We offer growth potential for motivated individuals, competitive compensation, and a comprehensive benefits package, including:

  • Medical, Dental, Vision, Life Insurance
  • 401K Retirement Plan
  • Paid Vacation Time
  • Paid Holidays
  • and More!
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Brio Water Technology • Glendale (CA)

On-site
USD 120,000 - 180,000
Data Engineer
Data Engineer

Home Organizers • Glendale (CA)

On-site
USD 200,000 - 250,000
Medical Insurance
Dental Insurance
Vision Insurance
+4
Data Engineer
Data Engineer

Brio Water Technology • Glendale (CA)

On-site
USD 200,000 - 250,000
401K Retirement Plan
Paid Vacation Time
Paid Holidays
+1
Data Engineer
Data Engineer

InfoVision Inc. • Detroit (MI)

On-site
USD 105,000 - 155,000
Data Engineer
Data Engineer

Function • United States

On-site
USD 120,000 - 170,000
401k Match
Health Insurance
Choose your own holidays
Data Engineer (Mid Level)-Orbit
Data Engineer (Mid Level)-Orbit

Irth Solutions • United States

Remote
USD 120,000 - 180,000
Medical Insurance
Flexible Work Options
401(k) Plan
+1
Senior Data Engineer (Databricks)
Senior Data Engineer (Databricks)

Rallyday Partners • Northern (KY)

On-site
USD 150,000 - 180,000
Lead Data Engineer
Lead Data Engineer

K2 Integrity • New York (NY)

On-site
USD 140,000 - 210,000
Staff Data Engineer - Gaming Industry (Latam)
Staff Data Engineer - Gaming Industry (Latam)

X-Team • United States

Remote
USD 120,000 - 180,000
Data Engineer
Data Engineer

InfoVision, Inc. • United States

On-site
USD 110,000 - 150,000