Senior Software Engineer - Data Engineering

Caterpillar Inc.

Bengaluru

On-site

INR 3,000,000 - 6,000,000

Full time

6 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Caterpillar Inc. in Bengaluru, Karnataka, seeks a highly motivated Data Engineer to join our data engineering team. The role emphasizes building scalable data pipelines on AWS and Snowflake, with Python and SQL as core skills.

You will design, develop, and optimize data workflows, collaborate with data scientists and business stakeholders, and integrate vector databases to enable AI workloads. Strong problem-solving and proactive mindset are essential.

Qualifications

  • 8+ years of experience in data engineering or related roles.
  • Strong hands‑on experience with AWS cloud services, including data and AI workloads.
  • Deep understanding of Snowflake architecture, performance tuning, and best practices.
  • Advanced proficiency in Python and SQL for data pipelines, transformations, and services.
  • Hands‑on experience with graph databases (e.g., Neo4j, Neptune) and vector databases (e.g., Milvus, Amazon OpenSearch).
  • Experience with version control systems (e.g., Git) and Git workflows.
  • Experience working with Azure DevOps (AzDO) boards for backlog management in Agile environments.
  • Excellent analytical and problem‑solving skills.
  • Strong communication and collaboration abilities.
  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.

Responsibilities

  • Design, develop, and maintain scalable data pipelines on AWS using services such as S3, Glue, Lambda, Redshift, and EMR.
  • Build and optimize data warehousing solutions using Snowflake, including performance tuning and data modeling.
  • Write efficient and reusable code in Python and SQL for data transformation and processing.
  • Collaborate with cross‑functional teams, including data scientists, analysts, and business stakeholders, to understand data requirements.
  • Integrate vector databases with LLM‑based applications and AI workflows.
  • Monitor, troubleshoot, and improve pipeline performance and reliability.
  • Ensure data quality, integrity, and security across all stages of the pipeline.
  • Participate in code reviews, architecture discussions, and continuous improvement initiatives.

Skills

AWS cloud services
Snowflake
Python
SQL
Graph databases
Neo4j
Milvus
Git
Azure DevOps
Data modeling
Data pipelines

Education

CS/Engineering degree

Tools

Amazon OpenSearch
Neptune

Job description

Job Description
Career Area

Technology, Digital and Data

Job Description

Your Work Shapes the World at Caterpillar Inc.

When you join Caterpillar, you'rejoining a global team who cares not just about the work we do – but also about each other. We are the makers, problem solvers, and future world builders who are creating stronger, more sustainable communities. We don'tjust talk about progress and innovation here – we make it happen, with our customers, where we work and live. Together, we are building a better world, so we can all enjoy living in it.

Job Summary

We are looking for a highly motivated and experienced Data Engineer to join our data engineering team. The ideal candidate will have a strong background in building scalable data pipelines using the AWS cloud stack and extensive hands‑on experience with Snowflake. Proficiency in Python and SQL, along with graph and vector database technologies, is essential. This role requires strong problem-solving abilities and a proactive mindset to deliver efficient, scalable, and reliable data solutions.

Key Responsibilities
  • Design, develop, and maintain scalable data pipelines on AWS using services such as S3, Glue, Lambda, Redshift, and EMR.
  • Build and optimize data warehousing solutions using Snowflake, including performance tuning and data modeling.
  • Write efficient and reusable code in Python and SQL for data transformation and processing.
  • Collaborate with cross‑functional teams, including data scientists, analysts, and business stakeholders, to understand data requirements.
  • Integrate vector databases with LLM‑based applications and AI workflows.
  • Monitor, troubleshoot, and improve pipeline performance and reliability.
  • Ensure data quality, integrity, and security across all stages of the pipeline.
  • Participate in code reviews, architecture discussions, and continuous improvement initiatives.
Required Qualifications
  • 8+ years of experience in data engineering or related roles.
  • Strong hands‑on experience with AWS cloud services, including data and AI workloads.
  • Deep understanding of Snowflake architecture, performance tuning, and best practices.
  • Advanced proficiency in Python and SQL for data pipelines, transformations, and services.
  • Hands‑on experience with graph databases (e.g., Neo4j, Neptune) and vector databases (e.g., Milvus, Amazon OpenSearch).
  • Experience with version control systems (e.g., Git) and Git workflows.
  • Experience working with Azure DevOps (AzDO) boards for backlog management in Agile environments.
  • Excellent analytical and problem‑solving skills.
  • Strong communication and collaboration abilities.
  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
Nice to Have skills
  • Knowledge of the NVIDIA ecosystem and its applications in data and AI.
  • Exposure to RAPIDS libraries (cuDF, cuML, cuGraph) or CUDA‑based tooling for GPU‑accelerated data processing, enabling faster transformation and optimization during large‑scale ingestion workflows.
  • Hands‑on expertise with vector databases, specifically Milvus, covering schema design, indexing, and optimizing write performance for large‑scale embedding ingestion pipelines.
  • Proficiency in building Knowledge Graph (Neo4J) ingestion pipelines using Graph Databases — including entity extraction, relationship modelling, and populating nodes and attributes.
Preferred Qualifications
  • Experience with orchestration tools such as AWS Step Functions.
  • Familiarity with data governance and compliance practices.
  • Exposure to real‑time data processing frameworks (e.g., Kafka, Spark Streaming).

This position requires working onsite five days a week.Relocation is available for this position.

Posting Dates

October 5, 2026 - October 6, 2026

Caterpillar is an Equal Opportunity Employer. Qualified applicants of any age are encouraged to apply

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Specialist
Senior Data Specialist

Caterpillar Financial Services Corporation • Bengaluru

On-site
INR 1,200,000 - 1,900,000
Senior Data Specialist
Senior Data Specialist

Caterpillar Brazil • Bengaluru

On-site
INR 1,800,000 - 2,600,000
Senior Software Engineer- Platform Services
Senior Software Engineer- Platform Services

Caterpillar Financial Services Corporation • Bengaluru

On-site
INR 4,000,000 - 6,000,000
Senior Data Engineer
Senior Data Engineer

Unison Group • Hyderabad

On-site
INR 800,000 - 1,500,000
Sr. Data Engineer
Sr. Data Engineer

Minfy Technologies • Gurugram District

On-site
INR 1,800,000 - 2,400,000
Senior Data Engineer
Senior Data Engineer

Proclink • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Data Engineer Snowflake
Data Engineer Snowflake

Digitrix Software LLP • India

On-site
INR 1,200,000 - 2,400,000
Senior Data Engineer
Senior Data Engineer

UST • Bengaluru

On-site
INR 1,800,000 - 3,000,000
Senior Data Engineer
Senior Data Engineer

Trantor • Dadri

On-site
INR 1,400,000 - 2,400,000
Data Engineer (Snowflake & AWS)
Data Engineer (Snowflake & AWS)

ContenTerra Software Pvt. Ltd. • Hyderabad

On-site
INR 1,200,000 - 1,800,000