Data Platform Engineer

Yotta

Delhi

On-site

INR 600,000 - 1,000,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Yotta is seeking a Junior Data Platform Engineer in India to support Apache Spark, Airflow, and JupyterHub environments with a strong Python and Linux foundation.

You will work with senior engineers to ensure smooth operation, deployment, and optimization of our data processing ecosystem and pipelines, building a career in big data engineering and data platform administration.

Qualifications

  • Bachelor’s or any relevant degree is required.
  • Strong foundation in Python and Linux.
  • Familiarity with Spark, Hive, Hadoop and Airflow.
  • Experience with Git, SQL and data pipelines preferred.
  • Exposure to Docker/Kubernetes and cloud platforms is a plus.

Responsibilities

  • Assist in setup, monitoring, and maintenance of Spark clusters, Hive, Hadoop and Airflow.
  • Support development and scheduling of data pipelines using Airflow DAGs and Python scripts.
  • Help manage JupyterHub for multi-user access and Spark integration.
  • Monitor cluster health and help troubleshoot Spark job failures.
  • Write Python automation scripts for data workflows and ETL tasks.

Skills

Python scripting
Linux
Apache Spark
Apache Hive
Hadoop
Airflow
Git
SQL
Docker
Kubernetes

Education

Bachelor’s or any relevant Degree

Tools

Docker
Kubernetes
JupyterHub
Spark Cluster
AWS/GCP/Azure

Job description

Yotta Data Services is India’s leading sovereign AIinfrastructure, cloud platform and data centre services company, enablingenterprises, governments, startups, and digital platforms to build, deploy,and scale next-generation AI and digital workloads securely within India.

With hyperscale data centre campuses in Navi Mumbai andGreater Noida (Delhi NCR), advanced GPU-powered AI infrastructure, and acomprehensive ecosystem of cloud, AI, hosting, cybersecurity, and managedplatform services, Yotta delivers high-performance, scalable, and compliantdigital infrastructure built for the AI era.

Yotta is at the forefront of powering India’s sovereign AIand digital transformation journey through world-class infrastructure, deeptechnology partnerships, and fully India-hosted enterprise-grade platforms.

Job Scope

We are looking for an enthusiastic Junior Data PlatformEngineer to support and manage our Apache Spark, Apache Airflow, andJupyterHub environments. This role is ideal for someone with a strongfoundation in Python and Linux, who is eager to build a career in big dataengineering and data platform administration.

You will work closely with senior engineers to ensuresmooth operation, deployment, and optimization of our data processingecosystem.

  • 1+years experience
Key Responsibilities
  • Assistin the setup, monitoring, and maintenance of Apache Spark clusters, ApacheHive, Hadoop and Airflow environments.
  • Supportthe development and scheduling of data pipelines using Airflow DAGs andPython scripts.
  • Helpmanage and configure JupyterHub for multi-user access and integration withSpark.
  • Monitorcluster health and performance under guidance and assist in troubleshootingSpark job failures.
  • Writeand maintain Python automation scripts for data workflows, ETL, and processautomation.
  • Participatein code reviews, documentation, and deployment activities.
  • Learnand follow best practices for distributed data processing, CI/CD, and DevOpsworkflows.
  • Collaboratewith senior engineers and data scientists to implement improvements and newfeatures.
Must-have skill
  • Basicunderstanding of Apache Spark, Apache Hive and Hadoop File System.
  • Familiaritywith Apache Airflow (understanding of DAGs, scheduling, and taskdependencies).
  • Hands-onexperience with Python scripting (data processing, automation, or APIinteraction).
  • Comfortableworking in Linux and container environments (command line, system logs,process management).
  • Goodunderstanding of data processing concepts, including ETL and distributedcomputing.
  • Basicknowledge of Git and version control.
Good-to-Have Skills
  • Exposureto Jupyter / JupyterHub for collaborative notebook environments.
  • Knowledgeof Docker or Kubernetes.
  • Knowledgeof Hadoop and Apache Spark Cluster.
  • Familiaritywith SQL and working with structured/unstructured data.
  • Experiencewith cloud platforms (AWS, GCP, or Azure) is a plus.
  • Interestin big data (Hadoop), DevOps, and data pipeline automation.
Qualifications Criteria
  • Bachelor’sor any relevant Degree.
Behavioral Attributes:
  • Artof skilful conversation
  • Creativity& Problem Solving
  • Learningon fly
  • BusinessAcumen
  • BuildingTrust
  • CustomerFocus
  • IntellectualHorsepower (Functional Skills)
  • ActionOrientation & Accountability
  • Listening,Sensing, Observing
  • DevelopingDirect Reports
  • Peripheral
  • BuildingCollaborative Relationships
Company Values
Customer Centricity
Agility
Integrity
Trust and Transparency
Happiness for all

Job Snapshot

Job ID

Job_751

Department

Operations, Service Delivery & CISO Function

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Cloud Engineer
Cloud Engineer

Yotta • India

On-site
INR 1,200,000 - 2,000,000
Data Engineer Pune · Hybrid Engineering · Full-time →
Data Engineer Pune · Hybrid Engineering · Full-time →

Woodfrog Tech OPC Private Limited • Pune District

Hybrid
INR 1,800,000 - 2,400,000
Hybrid working
Data Engineer
Data Engineer

Enterprise Minds • Pune District

Hybrid
INR 1,000,000 - 1,500,000
Big Data Developer
Big Data Developer

Viraaj HR Solutions Private Limited • Maharashtra

On-site
INR 1,000,000 - 1,500,000
Collaborative engineering culture
Competitive compensation
Training in cloud and Big Data technologies
Data Engineer Python - Senior Engineer
Data Engineer Python - Senior Engineer

Iris Software • Dadri

On-site
INR 1,800,000 - 2,400,000
Technical Writer
Technical Writer

Yotta • Delhi

On-site
INR 900,000 - 1,300,000
Data Engineer - Consultant
Data Engineer - Consultant

Iris Software • Dadri

On-site
INR 1,400,000 - 2,200,000
Data Engineer - SQL/PySpark
Data Engineer - SQL/PySpark

Forward Eye Technologies • Dadri

Hybrid
INR 1,400,000 - 2,000,000
Senior Data Engineer
Senior Data Engineer

Wissen • Bengaluru

Hybrid
INR 2,800,000 - 4,800,000
Senior Data Engineer
Senior Data Engineer

Wissen Technology • Bengaluru

Hybrid
INR 1,800,000 - 3,200,000