Job Title: Python Data Engineer
Experience: 3-6 Years
Location: Gurgaon (Work from Office)
Working Days: 5 Days a Week (Monday-Friday)
Employment Type: Full-Time
Compensation: Best in Industry
About the Role
We are looking for a highly skilled Python Data Engineer with a strong foundation in databases, SQL, and data processing pipelines. The ideal candidate should have hands‑on experience building scalable ETL/ELT solutions, working with large datasets, optimizing database performance, and developing reliable data pipelines on any major cloud platform (AWS, Azure, or GCP). This role requires someone who is passionate about writing clean, efficient Python code and has excellent database design and query optimization skills.
Key Responsibilities
- Design, develop, and maintain scalable data ingestion and processing pipelines.
- Develop ETL/ELT workflows for structured and semi‑structured data.
- Write efficient, maintainable, and production‑ready Python code.
- Design and optimize relational database schemas for high‑performance applications.
- Write complex SQL queries, stored procedures, views, and database functions.
- Perform database tuning, indexing, query optimization, and performance troubleshooting.
- Build automated data validation, transformation, and cleansing processes.
- Integrate multiple data sources including APIs, databases, files, and cloud storage.
- Work with large datasets while ensuring data quality, consistency, and reliability.
- Deploy and maintain data pipelines on cloud platforms (AWS/Azure/GCP).
- Collaborate with application developers, analysts, and business stakeholders.
- Troubleshoot production data issues and improve pipeline reliability.
- Follow software engineering best practices including version control, testing, and documentation.
Required Skills
Strong Python Skills
- Python 3.x
- Object‑Oriented Programming
- Pandas
- NumPy
- Logging
- Exception Handling
- Multi‑threading/Multiprocessing
- REST API Integration
Database Expertise (Mandatory)
Strong hands‑on experience with one or more of:
- PostgreSQL
- MySQL
- Microsoft SQL Server
- Oracle
Must have expertise in
- Advanced SQL
- Complex joins
- Window functions
- CTEs
- Query optimization
- Indexing strategies
- Execution plans
- Stored Procedures
- Views
- Database normalization
- Performance tuning
- Data modelling
- Transactions
- Constraints
Data Engineering
- ETL/ELT development
- Data transformation
- Batch processing
- Incremental data loads
- Data quality validation
- Data migration
- File processing (CSV, JSON, XML, Parquet)
- Data warehousing concepts
Cloud (Any One)
Experience with any one of the following:
- AWS
- Microsoft Azure
- Google Cloud Platform (GCP)
Exposure to cloud services such as object storage, managed databases, serverless compute, or data processing services is preferred.
Good to Have
- Apache Airflow
- Apache Spark (PySpark)
- Docker
- Kubernetes
- Git
- CI/CD pipelines
- Linux/Unix environment
- Message queues (Kafka/RabbitMQ)
- NoSQL databases (MongoDB, DynamoDB, Redis)
Qualifications
- Bachelor's or Master's degree in Computer Science, Information Technology, Engineering, or a related field.
- 3–6 years of hands‑on experience in Python development and data engineering.
What We Are Looking For
- Strong analytical and problem‑solving skills.
- Excellent SQL and database design expertise.
- Ability to work independently on end‑to‑end data engineering tasks.
- Experience handling large‑scale data processing and optimization.
- Good communication and collaboration skills.
- Passion for writing clean, efficient, and maintainable code.
Interview Process
- Technical Screening
- Technical Discussion
- Final Round (Face‑to‑Face in Gurgaon Office)
Note: Candidates should be willing to attend the final interview in person at our Gurgaon office.
If you have a passion for building robust data pipelines, solving complex database challenges, and working on cloud‑based data solutions, we'd love to hear from you.