Data Engineer / Data Architect

Soul Ai

Bengaluru

On-site

INR 1,200,000 - 1,800,000

Full time

6 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Soul Ai is hiring a Data Engineer / Data Architect to design, build, and maintain scalable data infrastructure for client ecosystems. The role emphasizes robust data pipelines, data modeling, and cross-functional collaboration with data science and engineering teams.

You will work across ETL/ELT processes, warehousing, lakehouse architecture, and cloud platforms, ensuring secure, reliable, and well-governed data workflows in production environments.

Qualifications

  • Strong programming skills in Python, SQL and Shell scripting.
  • Experience with ETL tools and orchestration frameworks (Airflow, Luigi, dbt).
  • Proficiency in relational and NoSQL databases (PostgreSQL, MySQL, MongoDB, Redis).
  • Experience with big data technologies: Spark, Kafka, Hive, Hadoop.
  • Understanding of data modeling, schema design and data warehousing concepts.
  • Cloud platforms (AWS/GCP/Azure) and services like Redshift, BigQuery, S3, Dataflow, or Databricks.
  • Knowledge of DevOps and CI/CD tools relevant to data infrastructure.

Responsibilities

  • Design and implement scalable, robust, and secure data pipelines.
  • Build ETL/ELT frameworks to collect, clean, and transform data (structured and unstructured).
  • Collaborate with data scientists, analysts, and backend engineers to enable data access and model integration.
  • Maintain data integrity, schema design, lineage, and quality monitoring.
  • Optimize performance and reliability of data workflows in production environments.
  • Design and manage data warehousing and lakehouse architecture.
  • Set up and manage IaC for infrastructure as needed.

Skills

Python
SQL
Shell scripting
Data pipelines
ETL concepts

Education

Bachelor's or Master's in Computer Science/Data Engineering/Information Systems

Tools

Airflow
Luigi
dbt
PostgreSQL
MySQL
MongoDB
Redis
Apache Spark
Kafka
Hive
Hadoop
Redshift
BigQuery
S3
Dataflow
Databricks
DevOps/CI-CD

Job description

We specialize in delivering high-quality human-curated data and AI-first scaled operations services

Based in San Francisco and Hyderabad, we are a fast-moving team on a mission to build AI for Good, driving innovation and societal impact

Role Overview:

We are seeking a Data Engineer / Data Architect who will be responsible for designing, building, and maintaining scalable data infrastructure and systems for a client

Youll play a key role in enabling efficient data flow, storage, transformation, and access across our organization or client ecosystems

Whether youre just beginning or already an expert, we value strong technical skills, curiosity, and the ability to translate complex requirements into reliable data pipelines

Responsibilities:
  • Design and implement scalable, robust, and secure data pipelines
  • Build ETL/ELT frameworks to collect, clean, and transform structured and unstructured data
  • Collaborate with data scientists, analysts, and backend engineers to enable seamless data access and model integration
  • Maintain data integrity, schema design, lineage, and quality monitoring
  • Optimize performance and ensure reliability of data workflows in production environments
  • Design and manage data warehousing and lakehouse architecture
  • Set up and manage infrastructure using IaC (Infrastructure as Code) when applicable
Required Skills:
  • Strong programming skills in Python, SQL, and Shell scripting
  • Hands-on experience with ETL tools and orchestration frameworks (e g, Airflow, Luigi, dbt)
  • Proficiency in relational databases (e g , PostgreSQL, MySQL) and NoSQL databases (e g, MongoDB, Redis)
  • Experience with big data technologies: Apache Spark, Kafka, Hive, Hadoop, etc
  • Deep understanding of data modeling, schema design, and data warehousing concepts
  • Proficient with cloud platforms (AWS/GCP/Azure) and services like Redshift, BigQuery, S3, Dataflow, or Databricks
  • Knowledge of DevOps and CI/CD tools relevant to data infrastructure
Nice to Have:
  • Experience working in real-time streaming environments
  • Familiarity with containerization and Kubernetes
  • Exposure to MLOps and collaboration with ML teams
  • Experience with security protocols, data governance, and compliance frameworks
Educational Qualifications:

Bachelors or Masters in Computer Science, Data Engineering, Information Systems, or a related technical field

Location - Mumbai, Delhi / NCR, Bengaluru , Kolkata, Chennai, Hyderabad, Ahmedabad, Pune, India

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer / Data Architect
Data Engineer / Data Architect

Soul Ai • Chennai District

On-site
INR 1,500,000 - 2,500,000
Data Engineer / Data Architect
Data Engineer / Data Architect

Soul Ai • New Delhi

On-site
INR 1,400,000 - 2,500,000
Data Engineer / Data Architect
Data Engineer / Data Architect

Soul Ai • Dadri

On-site
INR 1,200,000 - 1,800,000
Data Engineer / Data Architect
Data Engineer / Data Architect

Soul Ai • Mumbai

On-site
INR 1,500,000 - 3,000,000
Data Engineer / Data Architect
Data Engineer / Data Architect

Soul Ai • Gurugram District

On-site
INR 1,500,000 - 2,300,000
Data Engineer
Data Engineer

Agilisium • Chennai District

On-site
INR 800,000 - 1,200,000
Data Scientist
Data Scientist

Soul Ai • New Delhi

On-site
INR 1,200,000 - 2,400,000
Data Scientist
Data Scientist

Soul Ai • Mumbai

On-site
INR 1,200,000 - 1,800,000
Data Engineer
Data Engineer

e-Stone Information Technology Private Limited • Mumbai

On-site
INR 1,500,000 - 2,800,000
Data Scientist
Data Scientist

Soul Ai • Bengaluru

On-site
INR 1,200,000 - 2,400,000