Senior Data Engineer with Python (IR-491)

Intellectsoft Group

Bengaluru

Hybrid

INR 1,000,000 - 1,400,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
Life insurance
Accident insurance
Paid time off
Tech equipment provided
Udemy courses
Flexible hours

Job summary

Intellectsoft Group in Bengaluru, India, is seeking a data engineer to design and build scalable data pipelines using PySpark and modern data tooling. You will collaborate with data scientists to enhance model accuracy and performance, and create standardized data models across deployments.

The role requires strong Python, SQL, Spark, and Airflow skills, plus exposure to cloud platforms (AWS/Azure/GCP) and DevOps practices.

Qualifications

  • 5–6+ years of hands-on experience in Data Engineering.
  • Strong proficiency in Python and advanced SQL, including query optimization, data modeling, and performance tuning.
  • Deep understanding of distributed data processing frameworks, particularly Apache Spark.
  • Strong practical experience with Apache Airflow for workflow orchestration and pipeline management.
  • Working knowledge of backend application development frameworks such as FastAPI.
  • Foundational understanding of LLMs, AI/GenAI concepts, and their practical applications within data platforms.
  • Familiarity with modern AI and data infrastructure concepts, including: Vector Databases, Semantic Search, Knowledge Graphs, Retrieval-Augmented Generation (RAG architectures) (preferred).
  • Hands-on experience with at least one major cloud platform: AWS, Azure, or GCP.
  • Hands-on experience with Git and CI/CD pipelines.

Responsibilities

  • Design and build highly reliable and scalable data pipelines using PySpark and big data technologies.
  • Collaborate with the data science team to develop new features that enhance model accuracy and performance.
  • Create standardized data models to improve consistency across various deployments.
  • Troubleshoot and resolve issues in existing ETL pipelines and optimize workflows.
  • Conduct POCs to evaluate new technologies and integrate additional data sources.
  • Follow and promote best practices for software development, ensuring high-quality solutions that meet requirements and deadlines.
  • Document development updates and maintain clear technical documentation.

Skills

Python
SQL
Apache Spark
Apache Airflow
GenAI concepts
RAG
Vector Databases
Semantic Search
Knowledge Graphs
AWS
Azure
GCP
Git
CI/CD

Tools

Git
CI/CD

Job description

Intellectsoft is a software development company delivering innovative solutions since 2007. We operate across North America, Latin America, the Nordic region, the UK, and Europe. We specialize in industries like Fintech, Healthcare, EdTech, Construction, Hospitality, and more, partnering with startups, mid-sized businesses, and Fortune 500 companies to drive innovation and scalability. Our clients include Jaguar Motors, Universal Pictures, Harley-Davidson, and many more where our teams are making daily impactTogether, our team delivers solutions that make a difference.Learn more at www.intellectsoft.net

Our customer's product is an AI-powered platform that helps businesses make better decisions and work more efficiently. It uses advanced analytics and machine learning to analyze large amounts of data and provide useful insights and predictions. The platform is widely used in various industries, including healthcare, to optimize processes, improve customer experiences, and support innovation. It integrates easily with existing systems, making it easier for teams to make quick, data-driven decisions to deliver cutting-edge solutions.

  • 5–6+ years of hands-on experience in Data Engineering.
  • Strong proficiency in Python and advanced SQL, including query optimization, data modeling, and performance tuning.
  • Deep understanding of distributed data processing frameworks, particularly Apache Spark.
  • Strong practical experience with Apache Airflow for workflow orchestration and pipeline management.
  • Working knowledge of backend application development frameworks such as FastAPI.
  • Foundational understanding of LLMs, AI/GenAI concepts, and their practical applications within data platforms.
  • Familiarity with modern AI and data infrastructure concepts, including:
    • Vector Databases
    • Semantic Search
    • Knowledge Graphs
    • Retrieval-Augmented Generation (RAG) architectures (preferred)
  • Hands-on experience with at least one major cloud platform: AWS, Azure, or GCP.
  • Hands-on experience with Git and CI/CD pipelines.
Responsibilities
  • Design and build highly reliable and scalable data pipelines using PySpark and big data technologies.
  • Collaborate with the data science team to develop new features that enhance model accuracy and performance.
  • Create standardized data models to improve consistency across various deployments.
  • Troubleshoot and resolve issues in existing ETL pipelines and optimize workflows.
  • Conduct POCs to evaluate new technologies and integrate additional data sources.
  • Follow and promote best practices for software development, ensuring high-quality solutions that meet requirements and deadlines.
  • Document development updates and maintain clear technical documentation.
  • Employment-based cooperation
  • Comprehensive insurance for you and your family (health, life, and accident)
  • Paid PTO policy (vacation, sick leaves, and public holidays)
  • Tech equipment provided
  • Udemy courses, workshops, trainings & expert knowledge-sharing
  • Flexible hours & work setup
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer with Python (IR-491)
Senior Data Engineer with Python (IR-491)

Intellectsoft • Bengaluru

On-site
INR 800,000 - 1,200,000
Awesome projects with an impact
Udemy courses of your choice
Team-building events
+2
Senior Data Engineer
Senior Data Engineer

GlobalNodes • Gurgaon

On-site
INR 1,500,000 - 2,100,000
Data Engineer - Pyspark, Databricks, Snowflake, Azure Cloud
Data Engineer - Pyspark, Databricks, Snowflake, Azure Cloud

Optum India • Hyderabad

On-site
INR 1,500,000 - 2,700,000
Python/ETL Developer
Python/ETL Developer

Wissen • Bengaluru

Hybrid
INR 2,000,000 - 3,200,000
Senior Data Engineer
Senior Data Engineer

Wissen • Bengaluru

Hybrid
INR 2,800,000 - 4,800,000
Senior Data Engineer
Senior Data Engineer

Impronics Technologies • Indore

On-site
INR 1,200,000 - 1,800,000
Senior Fullstack Developer
Senior Fullstack Developer

agilisium • Chennai District

On-site
INR 800,000 - 1,200,000
Senior Software Engineer (Data)
Senior Software Engineer (Data)

KSOLVES • Indore District

On-site
Senior Data Engineer I
Senior Data Engineer I

RELX • Chennai District

On-site
INR 1,200,000 - 1,800,000
Senior Data Engineer
Senior Data Engineer

RBM Software • Pune District

On-site
INR 1,200,000 - 1,500,000
Opportunity to work on large-scale data projects
Exposure to modern cloud technologies
Collaborative work environment