Data Engineer AI & CLOUD Analytics

Hudson Data

Gurugram District

On-site

INR 1,500,000 - 2,100,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Hudson Data is seeking a Data Engineer with strong expertise in Python, SQL, GCP, and BigQuery to build scalable data platforms powering analytics, machine learning, and AI-driven products.

You will design reliable ETL/ELT pipelines, develop cloud-based data models, and collaborate with data scientists, analysts, and product teams to make high-quality data available for advanced analytics and AI use cases.

Qualifications

  • Experience building production-grade ETL/ELT pipelines.

Responsibilities

  • Design and maintain scalable ETL/ELT pipelines using Python, SQL, and GCP services.
  • Build optimized data models and analytics-ready datasets in BigQuery.
  • Integrate data from APIs, databases, files, and third-party platforms.
  • Support AI and machine-learning workflows through reliable feature and training datasets.
  • Implement data-quality checks, reconciliation, monitoring, logging, and error handling.
  • Optimize SQL queries, pipeline performance, and BigQuery cost efficiency.
  • Orchestrate batch and near-real-time workflows using Airflow or Cloud Composer.
  • Maintain data governance, security, lineage, and access controls.
  • Collaborate with data scientists, analysts, engineers, and business stakeholders.

Education

Bachelor or Master degree in Computer Science, Data Engineering, Data Science, Mathematics

Tools

Python
SQL
Google Cloud Platform
BigQuery
Airflow/Cloud Composer
DBT
Dataflow
Pub/Sub
Cloud Storage
Cloud Functions
Linux/Unix
Git

Job description

Role & responsibilities
About the Role

Hudson Data is seeking a Data Engineer with strong expertise in Python, SQL, GCP, and BigQuery to build scalable data platforms that power analytics, machine learning, and AI-driven products.

You will design reliable ETL/ELT pipelines, develop cloud-based data models, and work closely with data scientists, analysts, and product teams to make high-quality data available for advanced analytics and AI use cases.

Key Responsibilities
  • Design and maintain scalable ETL/ELT pipelines using Python, SQL, and GCP services.
  • Build optimized data models and analytics-ready datasets in BigQuery.
  • Integrate data from APIs, databases, files, and third-party platforms.
  • Support AI and machine-learning workflows through reliable feature and training datasets.
  • Implement data-quality checks, reconciliation, monitoring, logging, and error handling.
  • Optimize SQL queries, pipeline performance, and BigQuery cost efficiency.
  • Orchestrate batch and near-real-time workflows using Airflow or Cloud Composer.
  • Maintain data governance, security, lineage, and access controls.
  • Collaborate with data scientists, analysts, engineers, and business stakeholders.
Required Skills
  • Strong hands-on experience with Python for data processing and automation.
  • Advanced SQL, including complex joins, window functions, CTEs, query optimization, and performance tuning.
  • Strong experience with Google Cloud Platform.
  • Hands-on expertise in BigQuery, including data modeling, partitioning, clustering, and cost optimization.
  • Experience building production-grade ETL/ELT pipelines.
  • Familiarity with Airflow or Cloud Composer.
  • Understanding of data warehousing, star schema, and dimensional modeling.
  • Experience handling large datasets and implementing data-quality controls.
  • Working knowledge of Linux/Unix and Git.
AI-Oriented Experience
  • Experience preparing datasets for machine learning and generative AI applications.
  • Familiarity with feature engineering, model pipelines, and MLOps concepts.
  • Exposure to Vertex AI, embeddings, vector databases, or LLM-powered applications is preferred.
  • Understanding of model monitoring, retraining workflows, and responsible AI data practices is a plus.
Preferred Qualifications
  • Bachelors or Masters degree in Computer Science, Data Engineering, Data Science, Mathematics, or a related field.
  • Google Cloud certification, especially Professional Data Engineer, is preferred.
  • Experience with dbt, Dataflow, Pub/Sub, Cloud Storage, or Cloud Functions is an advantage.
Preferred candidate profile

At Hudson Data, you will work at the intersection of data engineering, cloud analytics, AI, and machine learning. You will contribute to consulting engagements and proprietary AI products while solving real-world business problems for global clients.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer (GCP)
Data Engineer (GCP)

Codvo Private Limited • Pune District

On-site
INR 1,000,000 - 2,000,000
Data Engineer
Data Engineer

Accenture in India • Bengaluru

On-site
INR 1,000,000 - 1,500,000
Data Engineer
Data Engineer

Advance Career Solutions • Pune District, Chennai District, Bengaluru

Hybrid
INR 1,200,000 - 2,800,000
Data Engineer
Data Engineer

Exillar • Ahmedabad District

On-site
INR 600,000 - 1,200,000
Growth opportunities
Structured career path
Culture of continuous learning
Data Engineer
Data Engineer

NARBA • Dadri

On-site
INR 600,000 - 900,000
Data Engineer
Data Engineer

Smart Ims • Bengaluru

On-site
INR 800,000 - 1,200,000
Senior Data Engineer
Senior Data Engineer

Expian Technologies • Bengaluru

Hybrid
INR 1,000,000 - 1,500,000
Data Platform - Senior Data Engineer
Data Platform - Senior Data Engineer

Renault Nissan Technology & Business Centre India • Chennai District

On-site
INR 1,500,000 - 2,500,000
Senior Data Engineer
Senior Data Engineer

Statusneo Technology Consulting • Gurugram District

On-site
INR 2,500,000 - 5,000,000
Data Engineer
Data Engineer

Recro • Bengaluru

On-site
INR 1,000,000 - 1,500,000