Senior Software Engineer - Ingestion

Databricks

Bengaluru

On-site

INR 3,500,000 - 7,000,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Databricks in Bengaluru is seeking a Lakeflow Connect software engineer to advance data ingestion for the Lakehouse platform, focusing on incremental data capture and log parsing to move data efficiently from various sources. You will integrate ingestion capabilities across surfaces and collaborate with backend and product teams to embed these capabilities into dashboards, notebooks, and SQL surfaces.

The role requires 5+ years of production experience in Python/Java/Scala/C++, experience with

Qualifications

  • BS (or higher) in Computer Science, or a related field.
  • 5+ years of production level experience in one of: Python, Java, Scala, C++, or similar language.
  • Experience developing large-scale distributed systems from scratch.
  • Experience in areas like Database replication, backup, transaction recovery at major database vendors is a plus.
  • Hands-on experience in developing and operating backend systems.
  • Ability to contribute effectively through all project phases from design to operations.

Responsibilities

  • Solve real business needs at scale by applying software engineering.
  • Deliver a highly scalable, available, and fault-tolerant engine processing hundreds of TB of data daily across thousands of customers.
  • Low level systems debugging, performance measurement and optimization on large production clusters.
  • Build architecture design, influence product roadmap, and take ownership over new projects.
  • Use deep experience to help prevent and investigate production issues.
  • Plan and lead complicated technical projects across multiple teams.

Skills

Python
Java
Scala
C++
Distributed systems

Education

BS in Computer Science

Job description

At Databricks, we are passionate about enabling data teams to solve the world's toughest problems - from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by building and running the world's best data and AI infrastructure platform so our customers can use deep data insights to improve their business.

Ingesting data into the Lakehouse is a strategic area of investment for Databricks and a key enabler for Data and AI workflows. Lakeflow Connect is looking to solve this problem by providing ready-to-use, point-and-click connectors for a wide variety of sources, including enterprise applications (like Salesforce, Workday, ServiceNow, SharePoint), databases (e.g., SQL Server), cloud storage, message queues, and local files.

In addition to being an important part of Lakeflow and Data Engineering, Connect is also a key platform capability. Every surface in Databricks (Dashboards, Notebooks, SQL, AI) requires ingestion capabilities and the lead for this role will need to work closely with other products to embed Connect into these surfaces.

We are looking for engineers with experience in core Database internals to join our Lakeflow Connect team. A key part of Connect is to extract data from OLTP systems while imposing minimal load on production systems. To do this efficiently we are building systems that use techniques such as incremental data capture, log parsing, etc. We are looking for engineers who continue to be hands‑on and are looking to make a large impact on an important problem for the company.

The Impact You Will Have
  • Solve real business needs at large scale by applying your software engineering.
  • Deliver a highly scalable, available, and fault-tolerant engine processing hundreds of TB of data daily across thousands of customers.
  • Low level systems debugging, performance measurement & optimization on large production clusters.
  • Build architecture design, influence product roadmap, and take ownership and responsibility over new projects.
  • Use your deep experience to help prevent and investigate production issues.
  • Plan and lead complicated technical projects that work with several teams within the company.
  • Break down complex problems quickly into potential solutions, knowns, and unknowns, and de‑risk (through prototyping/validation).
What We Look For
  • BS (or higher) in Computer Science, or a related field.
  • 5+ years of production level experience in one of: Python, Java, Scala, C++, or similar language.
  • Experience developing large-scale distributed systems from scratch.
  • Experience in areas like Database replication, backup, transaction recovery at one of the major database vendors (Microsoft SQL Server, Oracle, IBM etc.) is a plus to have.
  • Hands‑on experience in developing and operating backend systems.
  • Ability to contribute effectively throughout all project phases, from initial design and development to implementation and ongoing operations, with guidance from senior team members.
Benefits

At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region, please visit the benefits portal.

Our Commitment to Diversity and Inclusion

At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio‑economic status, veteran status, and other protected characteristics.

Compliance

If access to export‑controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer - Ingestion
Senior Software Engineer - Ingestion

Cacheflow • Bengaluru

On-site
INR 1,500,000 - 2,300,000
Staff Software Engineer - Ingestion
Staff Software Engineer - Ingestion

Cacheflow • Bengaluru

On-site
INR 4,500,000 - 9,000,000
Senior Software Engineer - Ingestion
Senior Software Engineer - Ingestion

Databricks Inc. • Bengaluru

On-site
INR 1,800,000 - 3,000,000
Engineering Manager (Ingestion)
Engineering Manager (Ingestion)

Databricks Inc. • Bengaluru

On-site
INR 3,500,000 - 8,000,000
Engineering Manager (Ingestion)
Engineering Manager (Ingestion)

Databricks • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Staff Software Engineer - Ingestion
Staff Software Engineer - Ingestion

Databricks • Bengaluru

On-site
INR 2,000,000 - 3,000,000
Comprehensive benefits
Diversity and inclusion commitment
Engineering Manager (Ingestion)
Engineering Manager (Ingestion)

Cacheflow • Bengaluru

On-site
INR 4,500,000 - 6,000,000
Staff Software Engineer - Data Platform
Staff Software Engineer - Data Platform

Menlo Ventures • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Comprehensive benefits
Diversity and inclusion programs
Staff Software Engineer (Data Platform) Databricks Bengaluru, India
Staff Software Engineer (Data Platform) Databricks Bengaluru, India

Neura Market • Bengaluru

On-site
INR 3,500,000 - 7,000,000
Senior Software Engineer (Data Platform)
Senior Software Engineer (Data Platform)

Cacheflow • Bengaluru

On-site
INR 3,500,000 - 6,000,000