Data Platform Lead

Mega Cloud Lab

Santa Clara (CA)

On-site

USD 150,000 - 190,000

Full time

4 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Mega Cloud Lab in Santa Clara, CA is seeking a Data Platform Lead who owns the technical direction and implementation of a secure, governed cloud data platform built on Databricks. The role combines hands-on data engineering with enterprise architecture leadership across tenant isolation and platform operations.

Responsibilities include defining target architecture, leading Databricks deployment, building data pipelines, ensuring security by design, data governance, encryption, and production

Qualifications

  • Bachelor’s degree or equivalent with substantial enterprise cloud data platform leadership experience.
  • Extensive hands-on Databricks, Spark, Delta Lake, Unity Catalog, SQL, and Python experience.
  • Proven ability to design and operate multi-tenant data platforms with governance and security considerations.

Responsibilities

  • Define target architecture and engineering standards for EPIC Databricks environments.
  • Lead Databricks setup including Unity Catalog, Delta Lake, pipelines, and workloads.
  • Design batch, streaming, and event-driven data ingestion and transformations.

Skills

Databricks
Apache Spark
Delta Lake
Unity Catalog
SQL
Python

Education

Bachelor's degree in Computer Science, Engineering, Information Systems, Data Science, or related field

Tools

Databricks platform

Job description

Data Platform Lead

Location: Santa Clara, CA

Function: Data Platforms / Advanced Analytics

Role Type: Technical Lead / Solution Lead

Primary Platform: Databricks Lakehouse on Cloud

Scope: EPIC Data Platform (Dedicated Tenant and Multi-Tenant Capabilities)

Role Purpose

The EPIC Data Platform Lead owns the technical direction and implementation leadership for a secure, governed, and scalable cloud data platform built on Databricks. This role combines hands-on data engineering and analytical problem-solving with enterprise architecture leadership across tenant isolation, data governance, identity and access, encryption, observability, production readiness, and platform operations.

The lead will translate business, security, and engineering requirements into implementable platform capabilities for internal, customer-dedicated, and controlled multi-tenant use cases.

Key Expected Outcomes

  • Trusted Data Products: Curated, traceable, and analytics-ready data with clear ownership, quality controls, and validation.

  • Secure Tenant Boundaries: Validated isolation across workspace, catalog, storage, identity, network, compute, jobs, APIs, and data exports.

  • Production-Grade Operations: Observable, supportable, and cost-aware services supported by automated deployment, runbooks, and evidence-based security controls.

Key Responsibilities

  • Platform Architecture & Technical Leadership: Define target architecture, engineering standards, roadmaps, Architecture Decision Records (ADRs), reusable patterns, and non-functional requirements for EPIC Databricks environments. Conduct design reviews and manage trade-offs across performance, security, operability, scalability, and cost.

  • Databricks Implementation: Lead the setup and optimization of workspaces, Unity Catalog, Delta Lake, data pipelines, Databricks Workflows, SQL Warehouses, compute policies, external locations, storage credentials, and deployment topologies. Establish maintainable Medallion Architecture (Bronze/Silver/Gold) and production practices.

  • Data Engineering & Analysis: Design batch, streaming, and event-driven ingestion, source-to-target transformations, reconciliation, data profiling, exploratory analysis, and root-cause analysis. Validate data latency, completeness, entity linking, accuracy, and business logic outcomes.

  • Security by Design: Partner with cybersecurity, IAM, network, and cloud teams to enforce least privilege, SSO/federation, service principal authentication, secrets management, private connectivity (PrivateLink), controlled egress, hardening, and vulnerability remediation.

  • Governance & Data Protection: Implement data classification, taxonomies, metadata management, automated lineage, retention rules, fine-grained access controls, row filters, column masks, controlled sharing, DLP-aligned controls, and evidence-driven compliance.

  • Encryption & Key Management: Design and implement encryption in transit and at rest, Customer-Managed Keys (CMK) / Bring Your Own Key (BYOK) patterns, cloud KMS/HSM integration, key separation, rotation, revocation, monitoring, recovery, and control validation.

  • Dedicated & Multi-Tenant Delivery: Define tenant onboarding, registry, provisioning, configuration, isolation, metadata-driven routing, metering/showback, offboarding, and migration. Prevent unauthorized cross-tenant access and validate isolation via automated negative testing.

  • Observability & Operations: Implement end-to-end logging, auditability, data-quality monitoring, health dashboards, alerting, SIEM integration, incident response runbooks, service-level measures, capacity planning, and cost management.

  • Delivery Leadership: Own backlog refinement, milestone delivery, risk mitigation, release readiness, production cutover, operational handoff, and cross-functional alignment. Mentor engineers and drive execution across data, cloud, security, QA, and business teams.

Required Qualifications

  • Education & Experience: Bachelor’s degree in Computer Science, Engineering, Information Systems, Data Science, or a related field (or equivalent practical experience) with substantial experience leading enterprise cloud data platform implementations.

  • Databricks Depth: Extensive hands-on technical proficiency with Databricks, Apache Spark, Delta Lake, Databricks Workflows, Unity Catalog, SQL, and Python (Scala is a plus).

  • Data Engineering & Analytics: Proven expertise in complex data analysis, profiling, reconciliation, debugging, performance tuning, and root-cause analysis on large-scale datasets.

  • Pipeline Design: Direct experience building production-grade batch, streaming, micro-batch, event, and file-based ingestion pipelines incorporating schema evolution, replayability, backfills, and idempotent processing.

  • Cloud Infrastructure: Strong hands-on knowledge of cloud-native data services, object storage, IAM, private networking, key management, logging/monitoring, IaC, and CI/CD pipelines (AWS preferred; Azure or GCP relevant).

  • Security & Governance: Deep understanding of least privilege, identity federation, service principals, data classification, lineage, retention, masking, audit logging, DLP, controlled sharing, and handling sensitive or regulated data.

  • Multi-Tenancy: Demonstrated experience designing or operating dedicated-tenant, multi-tenant, or customer-isolated platforms (tenant lifecycle, logical/physical isolation, resource governance, and cross-tenant testing).

  • Cryptographic Controls: Hands-on experience implementing encryption at rest/transit, cloud KMS/HSM integration, CMK/BYOK, key rotation, separation of duties, and audit evidence generation.

  • Leadership & Communication: Strong architecture, technical writing, stakeholder management, and team leadership skills to convert ambiguous requirements into executable engineering blueprints.

Preferred Qualifications

  • AWS & Databricks Integration: In-depth experience with Databricks on AWS (S3, KMS, PrivateLink, VPC Endpoints, IAM roles, CloudTrail, CloudWatch).

  • Streaming & Event Edge: Experience with Apache Kafka, streaming integration, or edge telemetry platforms.

  • DevSecOps Automation: Experience with Databricks Asset Bundles (DABs), Terraform, Git-based release pipelines, policy-as-code, automated testing, and environment promotion.

  • Advanced Databricks Workloads: Experience with Delta Sharing, MLflow, custom APIs, BI integration, AI/ML workloads, and governed data-product consumption patterns.

  • Security Frameworks: Familiarity with Zero Trust principles, NIST CSF, CIS Controls, threat modeling, penetration testing, and audit evidence practices.

  • Domain Experience: Experience in semiconductor manufacturing, R&D, lab/metrology data, equipment telemetry, or OT-integrated environments is a strong advantage.

  • Certifications: Relevant Databricks, Cloud Architecture, Data Engineering, or Security/Governance certifications.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Databricks SME
Databricks SME

Scicominfra • Atlanta (GA)

On-site
USD 180,000 - 240,000
Databricks Platform Engineer
Databricks Platform Engineer

EXL • New York (NY)

On-site
USD 150,000 - 190,000
Databricks Data Platform Architect
Databricks Data Platform Architect

UsefulBI Corporation • Raleigh (NC)

On-site
USD 130,000 - 180,000
Senior Data Engineer
Senior Data Engineer

Acestack • Irvine (CA)

On-site
USD 140,000 - 190,000
Databricks Engineer
Databricks Engineer

Tredence Inc. • Chicago (IL)

On-site
USD 140,000 - 190,000
Senior Platform Engineer (1149944)
Senior Platform Engineer (1149944)

The Judge Group • Tustin (CA)

On-site
USD 160,000 - 170,000
Databricks Practice Lead / Engineering Manager
Databricks Practice Lead / Engineering Manager

Scicominfra • Atlanta (GA)

On-site
USD 180,000 - 240,000
Data Team Lead
Data Team Lead

Incedo Inc. • San Rafael (CA)

Hybrid
USD 130,000 - 150,000
Medical insurance
Vision insurance
401(k)
Data Solution Architect
Data Solution Architect

Falcon Smart IT (FalconSmartIT) • San Mateo (CA)

On-site
USD 180,000 - 240,000
Senior Data Engineer
Senior Data Engineer

Prosum • Glendale (CA)

On-site
USD 150,000 - 210,000