Senior Data Engineer: Spark Pipelines & Production APIs

Keka Technologies Private Limited

United States

On-site

USD 140,000 - 190,000

Full time

35 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Keka Technologies Private Limited is seeking a senior data engineer to build, optimize, and operate production Spark pipelines that generate attribution feeds from large-scale data. You will design data processing across modern data lake and warehouse tech, orchestrated with AWS workflows.

The role involves building partner integrations, maintaining REST APIs (PHP/Symfony), diagnosing failures, backfills, and delivering automated tests.

Qualifications

  • 7+ years of professional software or data engineering experience.
  • Advanced SQL, relational data modeling, and large-scale data processing experience.

Responsibilities

  • Build, optimize, and operate production Apache Spark pipelines for attributed conversion feeds.
  • Design and evolve data processing across data lake/warehouse with AWS-based orchestration.
  • Develop and operate REST APIs, including PHP-based APIs, with secure, backward-compatible changes.
  • Diagnose production failures, backfills, and manage staged releases and rollbacks.
  • Write unit and integration tests for pipelines, APIs, and feeds.

Skills

Java
Python
SQL
Spark

Tools

Snowflake
MySQL
Iceberg
Airflow
Symfony
PHP

Job description

Keka Technologies Private Limited is seeking a senior data engineer to build, optimize, and operate production Spark pipelines that generate attribution feeds from large-scale data. You will design data processing across modern data lake and warehouse tech, orchestrated with AWS workflows.

The role involves building partner integrations, maintaining REST APIs (PHP/Symfony), diagnosing failures, backfills, and delivering automated tests.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Engineer — Databricks & Spark Pipelines
Senior Data Engineer — Databricks & Spark Pipelines

Prosum • Glendale (CA)

On-site
USD 150,000 - 210,000
Senior Data Engineer: Spark & Cloud Pipelines
Senior Data Engineer: Spark & Cloud Pipelines

Tata Consultancy Services • Irving (TX)

On-site
USD 90,000 - 110,000
Discretionary Annual Incentive.
Comprehensive Medical Coverage
401K Plan
+1
Senior Backend Engineer: Spark, Python, Scala on AWS/GCP
Senior Backend Engineer: Spark, Python, Scala on AWS/GCP

MACHINE LEARNING TECHNOLOGIES LLC • Elk Grove (CA)

On-site
USD 120,000 - 160,000
Senior Data Engineer: Pipelines, Cloud & Data Lakes
Senior Data Engineer: Pipelines, Cloud & Data Lakes

hackajob • Plano (TX)

On-site
USD 110,000 - 150,000
Health care coverage
On-site wellness centers
Retirement plan
+4
Remote Senior Data Engineer: Gen AI Powered Pipelines
Remote Senior Data Engineer: Gen AI Powered Pipelines

Resonate • Northern (KY)

Hybrid
USD 120,000 - 180,000
401(k) match
Open PTO
Remote-first environment
Senior Data Engineer: Build Scalable Pipelines & Data Lakes
Senior Data Engineer: Build Scalable Pipelines & Data Lakes

XPEL, Inc. • San Antonio (TX)

On-site
USD 110,000 - 150,000
Senior Data Engineer: Scale Pipelines & Lakehouse Leadership
Senior Data Engineer: Scale Pipelines & Lakehouse Leadership

Survey Sampling International Hyderabad Private Ltd. (India) • Westport (CT)

On-site
USD 130,000 - 150,000
Medical benefits
Discretionary incentive program
Senior Data Engineer: Scalable Spark Pipelines & Real-Time
Senior Data Engineer: Scalable Spark Pipelines & Real-Time

3M Consultancy • Washington

On-site
USD 100,000 - 130,000
Senior Data Engineer: Lead Lakehouse Pipelines & ETL
Senior Data Engineer: Lead Lakehouse Pipelines & ETL

Dynata • United States

On-site
USD 130,000 - 150,000
Medical benefits
Discretionary incentive program
Data Analyst
Data Analyst

Keka Technologies Private Limited • United States

On-site
USD 140,000 - 190,000