Senior Data Platform Engineer

DigiCert

Bengaluru

On-site

INR 4,000,000 - 7,000,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

DigiCert is seeking a Senior Data Platform Engineer to design, build, and operate the foundational data and ML infrastructure that powers analytics and machine learning across the organization. You will own Databricks-based pipelines, data models, and governance to enable scalable, reliable analytics.

You will collaborate with data science, analytics, and product teams to deliver a platform that is performant, observable, and cost-aware, while mentoring junior engineers and aligning with

Qualifications

  • 6+ years of experience in data engineering, data platform, or ML engineering roles
  • Strong proficiency in Python and SQL, with a track record of building production-grade data pipelines using both
  • Hands-on Databricks expertise: Delta Lake, Unity Catalog, Databricks Workflows, PySpark, and the Databricks ecosystem broadly
  • Experience building and maintaining ML pipelines in production — feature engineering, training pipelines, experiment tracking, and model deployment
  • Familiarity with MLflow or comparable experiment tracking and model registry tools
  • Experience working on cloud data platforms (AWS, Azure, or GCP)
  • Strong understanding of data modeling, dimensional design, and analytics-friendly data architecture
  • Experience with batch and incremental/CDC pipeline patterns
  • Proficiency with Git, version control, and CI/CD practices for data and ML workflows
  • Strong engineering judgment — you think about reliability, maintainability, and cost, not just correctness
  • Clear communication and comfort working with both technical and non-technical stakeholders

Responsibilities

  • Design, build, and maintain scalable data and ML pipelines using Python and SQL, processing large-scale datasets across batch and streaming workloads
  • Own and evolve core platform infrastructure on Databricks — including Delta Lake table architecture, Unity Catalog governance, Databricks Workflows orchestration, and compute optimization
  • Build and maintain end-to-end ML pipelines: feature engineering, model training pipelines, experiment tracking (MLflow), and model deployment/serving infrastructure
  • Collaborate with data scientists to operationalize models — bridging the gap between experimentation and production‑grade ML systems
  • Define and enforce data platform standards: ingestion patterns, data modeling conventions, medallion architecture (Bronze/Silver/Gold), and pipeline reliability practices
  • Implement data quality, observability, and monitoring frameworks to ensure platform health and data trustworthiness
  • Optimize pipelines for performance, cost, and reliability at scale using Spark and PySpark
  • Evaluate, integrate, and govern new platform tooling and data sources within the Databricks ecosystem
  • Contribute to architectural decisions and help drive the long‑term data platform roadmap
  • Participate in code reviews, technical design discussions, and engineering standards
  • Mentor junior engineers and elevate overall platform and data engineering practices
  • Document platform architecture, pipeline design, and operational runbooks

Skills

Python
SQL
Databricks
ML pipelines
Cloud platforms

Tools

Git
CI/CD
Spark

Job description

DigiCert is a global leader in intelligent trust. We protect the digital world by ensuring the security, privacy, and authenticity of every interaction. Our AI-powered DigiCert ONE platform unifies PKI, DNS, and certificate lifecycle management, to secure infrastructure, software, devices, messages, AI content and agents. Learn why more than 100,000 organizations, including 90% of the Fortune 500, choose DigiCert to stop today’s threats and prepare for a quantum-safe future atwww.digicert.com

Job summary

We are looking for a Senior Data Platform Engineer to design, build, and operate the foundational data and ML infrastructure that powers analytics, reporting, and machine learning across DigiCert. This role sits at the intersection of data engineering and ML platform work — you will own the systems, pipelines, and tooling that data scientists, analysts, and engineers rely on every day. You bring deep expertise in Databricks, a strong engineering mindset, and hands‑on experience building and maintaining ML pipelines in production. You will partner closely with data science, analytics, product, and business teams to deliver a platform that is reliable, governed, and built to scale.

What you will do
  • Design, build, and maintain scalable data and ML pipelines using Python and SQL, processing large-scale datasets across batch and streaming workloads
  • Own and evolve core platform infrastructure on Databricks — including Delta Lake table architecture, Unity Catalog governance, Databricks Workflows orchestration, and compute optimization
  • Build and maintain end-to-end ML pipelines: feature engineering, model training pipelines, experiment tracking (MLflow), and model deployment/serving infrastructure
  • Collaborate with data scientists to operationalize models — bridging the gap between experimentation and production‑grade ML systems
  • Define and enforce data platform standards: ingestion patterns, data modeling conventions, medallion architecture (Bronze/Silver/Gold), and pipeline reliability practices
  • Implement data quality, observability, and monitoring frameworks to ensure platform health and data trustworthiness
  • Optimize pipelines for performance, cost, and reliability at scale using Spark and PySpark
  • Evaluate, integrate, and govern new platform tooling and data sources within the Databricks ecosystem
  • Contribute to architectural decisions and help drive the long‑term data platform roadmap
  • Participate in code reviews, technical design discussions, and engineering standards
  • Mentor junior engineers and elevate overall platform and data engineering practices
  • Document platform architecture, pipeline design, and operational runbooks
What you will have
  • 6+ years of experience in data engineering, data platform, or ML engineering roles
  • Strong proficiency in Python and SQL, with a track record of building production‑grade data pipelines using both
  • Hands‑on Databricks expertise: Delta Lake, Unity Catalog, Databricks Workflows, PySpark, and the Databricks ecosystem broadly
  • Experience building and maintaining ML pipelines in production — feature engineering, training pipelines, experiment tracking, and model deployment
  • Familiarity with MLflow or comparable experiment tracking and model registry tools
  • Experience working on cloud data platforms (AWS, Azure, or GCP)
  • Strong understanding of data modeling, dimensional design, and analytics‑friendly data architecture
  • Experience with batch and incremental/CDC pipeline patterns
  • Proficiency with Git, version control, and CI/CD practices for data and ML workflows
  • Strong engineering judgment — you think about reliability, maintainability, and cost, not just correctness
  • Clear communication and comfort working with both technical and non‑technical stakeholders
Nice to have
  • Experience with streaming or near real‑time pipelines (Kafka, Kinesis, Spark Structured Streaming)
  • Familiarity with feature store platforms (Databricks Feature Store, Feast, or Tecton)
  • Experience with LLM pipelines, RAG architectures, or AI/BI tooling (Genie, AI Functions)
  • Knowledge of data quality and observability tooling (Great Expectations, Monte Carlo, etc.)
  • Exposure to dbt or similar SQL‑based transformation frameworks
  • Infrastructure‑as‑code experience (Terraform, Databricks Asset Bundles)
  • Experience working in Agile or Scrum environments
  • Prior experience mentoring engineers or shaping platform standards
  • Generous time off policies
  • Education, wellness and lifestyle support
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AWS Databricks Platform Engineer
AWS Databricks Platform Engineer

Qualcomm • Hyderabad

On-site
INR 2,300,000 - 3,000,000
Databricks Platform Engineer
Databricks Platform Engineer

Navlakha Management Services • Pune District, Thiruvananthapuram

Hybrid
INR 2,500,000 - 4,000,000
Principal Engineer – Data Platforms & MLOps (Databricks)
Principal Engineer – Data Platforms & MLOps (Databricks)

Codvo Private Limited • Pune District

On-site
INR 2,000,000 - 3,000,000
Senior Data Engineer
Senior Data Engineer

GlobalNodes • Gurgaon

On-site
INR 1,500,000 - 2,100,000
Databricks - Data Engineer
Databricks - Data Engineer

Tredence Inc. • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Senior Databricks Engineer
Senior Databricks Engineer

DataBeat • Hyderabad

On-site
INR 2,500,000 - 4,200,000
Data Engineer - Lead
Data Engineer - Lead

Iris Software • Dadri

On-site
INR 1,500,000 - 2,500,000
Senior Data Platform Engineer – ModelOps
Senior Data Platform Engineer – ModelOps

Digi-Key Electronics • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Senior Data Platform Engineer
Senior Data Platform Engineer

Digi-Key Electronics • India

On-site
INR 1,500,000 - 2,500,000
Data Platform Engineer
Data Platform Engineer

Purview Services • Hyderabad, Pune District

Hybrid
INR 2,500,000 - 5,200,000