Data Engineer / Data Architect

Soul Ai

Gurugram District

On-site

INR 1,500,000 - 2,300,000

Full time

6 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Soul Ai seeks a Data Engineer / Data Architect to design, build, and maintain scalable data infrastructure for client ecosystems. You will enable efficient data flow, storage, transformation, and access across teams, translating complex requirements into reliable data pipelines.

We value strong technical skills and curiosity, with emphasis on robust data modeling, governance and production workflows. Local collaboration on client projects is expected.

Qualifications

  • Proficient in Python, SQL and shell scripting for data workflows.
  • Hands-on experience with ETL/ELT tools and orchestration (Airflow, Luigi, dbt).
  • Strong in relational and NoSQL databases (PostgreSQL, MySQL, MongoDB, Redis).
  • Experience with big data tech: Spark, Kafka, Hive, Hadoop.
  • Solid data modeling, schema design, and warehousing concepts.
  • Cloud expertise across AWS/GCP/Azure and services like Redshift, BigQuery, S3, Dataflow, Databricks.
  • Familiar with DevOps/CI/CD practices for data platforms.

Responsibilities

  • Design scalable, robust, secure data pipelines.
  • Build ETL/ELT frameworks for structured and unstructured data.
  • Collaborate with data scientists, analysts and engineers to enable data access and model integration.
  • Maintain data integrity, schemas, lineage and quality monitoring.
  • Optimize production data workflows for performance and reliability.
  • Design and manage data warehousing and lakehouse architectures.
  • Set up IaC where applicable to automate infrastructure.

Skills

Python
SQL
Shell scripting
Data modeling
Data warehousing
CI/CD for data infra

Education

Bachelor's or Master's in CS/Data Engineering/IS

Tools

Airflow
Luigi
dbt
PostgreSQL
MySQL
MongoDB
Redis
Apache Spark
Kafka
Hive
Hadoop
Redshift
BigQuery
S3
Dataflow
Databricks

Job description

We specialize in delivering high-quality human-curated data and AI-first scaled operations services

Based in San Francisco and Hyderabad, we are a fast-moving team on a mission to build AI for Good, driving innovation and societal impact

Role Overview:

We are seeking a Data Engineer / Data Architect who will be responsible for designing, building, and maintaining scalable data infrastructure and systems for a client

Youll play a key role in enabling efficient data flow, storage, transformation, and access across our organization or client ecosystems

Whether youre just beginning or already an expert, we value strong technical skills, curiosity, and the ability to translate complex requirements into reliable data pipelines

Responsibilities:
  • Design and implement scalable, robust, and secure data pipelines
  • Build ETL/ELT frameworks to collect, clean, and transform structured and unstructured data
  • Collaborate with data scientists, analysts, and backend engineers to enable seamless data access and model integration
  • Maintain data integrity, schema design, lineage, and quality monitoring
  • Optimize performance and ensure reliability of data workflows in production environments
  • Design and manage data warehousing and lakehouse architecture
  • Set up and manage infrastructure using IaC (Infrastructure as Code) when applicable
Required Skills:
  • Strong programming skills in Python, SQL, and Shell scripting
  • Hands-on experience with ETL tools and orchestration frameworks (e g, Airflow, Luigi, dbt)
  • Proficiency in relational databases (e g , PostgreSQL, MySQL) and NoSQL databases (e g , MongoDB, Redis)
  • Experience with big data technologies: Apache Spark, Kafka, Hive, Hadoop, etc
  • Deep understanding of data modeling, schema design, and data warehousing concepts
  • Proficient with cloud platforms (AWS/GCP/Azure) and services like Redshift, BigQuery, S3, Dataflow, or Databricks
  • Knowledge of DevOps and CI/CD tools relevant to data infrastructure
Nice to Have:
  • Experience working in real-time streaming environments
  • Familiarity with containerization and Kubernetes
  • Exposure to MLOps and collaboration with ML teams
  • Experience with security protocols, data governance, and compliance frameworks
Educational Qualifications:

Bachelors or Masters in Computer Science, Data Engineering, Information Systems, or a related technical field

Location - Mumbai, Delhi / NCR, Bengaluru , Kolkata, Chennai, Hyderabad, Ahmedabad, Pune, India

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer / Data Architect
Data Engineer / Data Architect

Soul Ai • Mumbai

On-site
INR 1,500,000 - 3,000,000
Data Engineer / Data Architect
Data Engineer / Data Architect

Soul Ai • Chennai District

On-site
INR 1,500,000 - 2,500,000
Data Engineer / Data Architect
Data Engineer / Data Architect

Soul Ai • New Delhi

On-site
INR 1,400,000 - 2,500,000
Data Engineer / Data Architect
Data Engineer / Data Architect

Soul Ai • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Data Engineer / Data Architect
Data Engineer / Data Architect

Soul Ai • Dadri

On-site
INR 1,200,000 - 1,800,000
Data Engineer - Gurugram
Data Engineer - Gurugram

Yeah! Global • Gurugram District

On-site
INR 1,200,000 - 1,800,000
Data Engineer
Data Engineer

Agilisium • Chennai District

On-site
INR 800,000 - 1,200,000
Data Engineer - Bangalore
Data Engineer - Bangalore

Yeah! Global • Bengaluru

On-site
INR 800,000 - 1,200,000
Data Scientist
Data Scientist

Soul Ai • Bengaluru

On-site
INR 1,200,000 - 2,400,000
Data Scientist
Data Scientist

Soul Ai • New Delhi

On-site
INR 1,200,000 - 2,400,000