Python Data Extraction Engineer

Zohorecruit

Coimbatore District

On-site

INR 900,000 - 1,300,000

Full time

11 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

VSERVE EBUSINESS SOLUTIONS INDIA PRIVATE LIMITED seeks a Python Data Extraction Engineer in Coimbatore to build robust web-scraping pipelines for government and public sources. You will create crawlers, extract structured data from HTML pages and APIs, and ensure data quality across multiple portals.

The role requires 3–7 years of experience with Python, pandas, web scraping, and automation using Selenium/Playwright. You will collaborate with GIS and engineering teams to enhance data coverage.

Qualifications

  • Strong hands-on experience with Python and web scraping.
  • Familiar with HTML parsing, REST APIs and data cleaning.
  • Experience using Selenium or Playwright for automation.

Responsibilities

  • Identify and evaluate government/public data sources relevant to property, businesses and activity.
  • Build Python-based crawlers and extraction pipelines for government portals and public websites.
  • Extract structured information from HTML pages, tables, PDFs, and public APIs.
  • Automate recurring extraction from multiple sources while respecting access rules and rate limits.
  • Clean, standardize and normalize data from different authorities.
  • Perform entity matching across datasets using key fields.
  • Develop mechanisms to detect schema changes and extraction failures.
  • Store extracted data in structured databases for downstream use.

Skills

Python
Web scraping
HTTP clients
BeautifulSoup/lxml
Selenium/Playwright
REST APIs/JSON
SQL
Data cleaning
Regex
Git

Tools

Selenium
Playwright
BeautifulSoup
lxml
Requests
Git

Job description

VSERVE EBUSINESS SOLUTIONS INDIA PRIVATE LIMITED | Full time

Python Data Extraction Engineer

Coimbatore, India | Posted on 09/28/2026

  • Industry E-Commerce/Supply Chain Management
  • Date Opened 09/28/2026
  • Job Type Full time
  • City Coimbatore
  • State/Province Tamil Nadu
  • Country India
About Us

Vserve Ebusiness Solution is one of the leading providers of e-commerce solutions and product catalog management services. We provide best-in-class services at affordable rates. We can build your brand like ours. With a track record of handling various industries in the US, UK, Europe, and Asia-Pacific, Vserve’s highly experienced team is dedicated to delivering the finest e-commerce solutions with the highest quality of service. With Vserve, you will receive customer satisfaction as we strive to exceed expectations. All in all, Vserve is the ideal partner for you to push your business forward.

erve Ebusiness Solution is one of the leading providers of e-commerce solutions and product catalog management services. We provide best-in-class services at affordable rates. We can build your brand like ours. With a track record of handling various industries in the US, UK, Europe, and Asia-Pacific, Vserve’s highly experienced team is dedicated to delivering the finest e-commerce solutions with the highest quality of service. With Vserve, you will receive customer satisfaction as we strive to exceed expectations. All in all, Vserve is the ideal partner for you to push your business forward.

Job Description

Python Data Extraction Engineer - Web Scraping & Government Data
Location:Coimbatore
Experience:3-7 years
Role Type:Full-time

Identify and evaluate government and public data sources relevant to property, businesses and commercial activity.

Build Python-based crawlers and extraction pipelines for government portals and public websites.

Extract structured information from HTML pages, tables, PDFs, downloadable files and publicly accessible APIs.

Automate recurring extraction from multiple sources while respecting applicable access rules, rate limits, and terms.

Clean, standardize and normalize inconsistent data from different government authorities.

Perform entity matching and record linkage across datasets using fields such as owner name, company name, address, survey number, coordinates and project information.

Develop mechanisms to detect website/schema changes and extraction failures.

Build validation and QA processes to measure completeness and accuracy.

Store extracted information in structured databases and expose clean datasets to downstream applications.

Work closely with GIS, product and engineering teams to combine location-based signals with government/public records.

Research new public and open-data sources that can improve the accuracy and completeness of our intelligence.

Examples of Data Sources The work may involve sources such as:

  • State planning and development authorities
  • DTCP and similar planning authorities
  • RERA databases
  • Building and planning permission records
  • Land and property records
  • Tender and procurement portals
  • Company/business registries
  • Government open-data portals
  • Environmental and regulatory approvals
  • Public notices and downloadable government documents
  • Maps and geospatial datasets
  • Other legally accessible public and open-source information

Required Technical Skills Strong hands-on experience with:

  • Python
  • Web scraping and crawling
  • Requests / HTTP clients
  • BeautifulSoup / lxml
  • Selenium and/or Playwright
  • REST APIs and JSON
  • HTML/XML parsing
  • SQL
  • Data cleaning and transformation
  • Regex and text processing
  • Git

What We Are Looking For We particularly want someone who is a problem solver rather than simply a Python programmer . The candidate should be able to investigate the available sources, understand how the underlying website works, determine the best extraction approach, build the pipeline and validate the resulting data. Ideal Background Candidates may come from backgrounds such as:

  • Web scraping / data extraction companies
  • Alternative-data companies
  • PropTech / real-estate data companies
  • Market-intelligence companies
  • OSINT/data intelligence companies
  • Government-data projects
  • Lead/data enrichment companies
  • GIS/location-intelligence companies

Success in This Role Within the first few months, the successful candidate should be able to:

  • Map relevant government/public data sources.
  • Build reliable extraction pipelines across multiple portals.
  • Convert fragmented information into standardized records.
  • Cross-reference records from multiple sources.
  • Establish automated QA and monitoring.
  • Continuously discover additional datasets that improve our product's coverage and accuracy.

The objective is not simply to scrape websites. It is to build a scalable public-data acquisition and enrichment engine that becomes a core component of our intelligence platform.

Requirements

Python, pandas, webscraping, web crawling,selenium postman

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Python Data Extraction Engineer
Python Data Extraction Engineer

Vserve • Coimbatore District

On-site
INR 900,000 - 1,300,000
Senior Data Engineer (Web Scraping)
Senior Data Engineer (Web Scraping)

Oxford Data Plan Ltd. • Chennai District

On-site
INR 2,500,000 - 4,000,000
Senior Python Developer - Web Scraping & Data Crawling
Senior Python Developer - Web Scraping & Data Crawling

Globaldata • Hyderabad

On-site
INR 1,000,000 - 2,000,000
Software Engineer, Data Infrastructure and Acquisition
Software Engineer, Data Infrastructure and Acquisition

Analogy Group • India

On-site
INR 3,500,000 - 6,000,000
Senior Data Engineer Web Scraping
Senior Data Engineer Web Scraping

Oxford Data Plan • Indore District

On-site
INR 1,500,000 - 2,100,000
Project Coordinator – Web Scraping & Data Intelligence
Project Coordinator – Web Scraping & Data Intelligence

Actowiz Solutions • Ahmedabad District

On-site
INR 600,000 - 1,200,000
Python Developer
Python Developer

Ascendion • Gurugram District

On-site
INR 1,000,000 - 2,000,000
Senior Data Engineer
Senior Data Engineer

Foss United • New Delhi

On-site
INR 900,000 - 1,500,000
SME-Web Scraping and Crawling — Web Data Platform
SME-Web Scraping and Crawling — Web Data Platform

Three Across • Hyderabad

On-site
INR 2,500,000 - 4,500,000
Python Mid/Senior Developer – Web Scraping & Automation
Python Mid/Senior Developer – Web Scraping & Automation

Hr Actowiz Solutions • Ahmedabad District

On-site
INR 1,400,000 - 2,100,000