A data engineering company based in India is seeking a skilled data engineer to design and maintain scalable ETL/ELT pipelines. The ideal candidate will have strong proficiency in Python and SQL, experience with PostgreSQL and SQLite, and familiarity with distributed data processing tools like Spark and Airflow. This is a flexible, hourly contract position ideal for self-motivated individuals looking to work remotely.
Qualifications
Proficiency in Python, SQL, and data engineering techniques.
Hands-on experience with relational databases like PostgreSQL and SQLite.
Familiarity with distributed data processing concepts and tools.
Responsibilities
Design, build, and maintain scalable ETL/ELT pipelines.
Validate and serve datasets for analytics and machine learning.
Optimize pipeline performance using Python and SQL.
Skills
Python
SQL
Data engineering
PostgreSQL
SQLite
Airflow
Pandas
Spark
Education
Background in computer science or data engineering
Tools
Airflow
Spark
DuckDB
Job description
Base pay range
$14.00/hr - $14.00/hr
Type: Hourly contract
Compensation: $14/hour
Location: India
Commitment: 10-40 hours/week, flexible and asynchronous
Role Responsibilities
Design, build, and maintain scalable ETL/ELT pipelines to support analytics and machine learning workflows.
Validate, enrich, and serve datasets to ensure reliability, reproducibility, and ML readiness.
Define and manage schemas, data contracts, and versioning with strong discipline.
Work with relational databases such as PostgreSQL and SQLite.
Implement distributed data processing using tools like Spark or DuckDB.
Orchestrate and monitor workflows using Airflow or similar scheduling tools.
Optimize pipeline performance and resilience using Python, pandas, and SQL.
Collaborate with researchers and engineers to align data pipelines with AI research and production needs.
Requirements
Background in computer science, data engineering, or information systems.
Proficiency in Python, pandas, and SQL.
Hands‑on experience with PostgreSQL, SQLite, or similar databases.
Understanding of distributed data processing frameworks such as Spark or DuckDB.
Experience orchestrating workflows with Airflow or comparable tools.
Familiarity with common data formats including JSON, CSV, and Parquet.
Strong focus on schema design, data quality, and version control using Git.
Ability to work independently in a remote, project‑based environment.
Application Process (Takes 20 Min)
Upload resume
Interview (15 min)
Seniority level
Associate
Employment type
Contract
Job function
Information Technology
Software Development, IT Services and IT Consulting, and Research Services