1,229,538open jobs
69,967companies
213,807added this week
Browse all
Location
In office (Coimbatore)
Seniority
Middle · 3+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 4, 2026. First seen by Alion on Oct 3, 2026.

Overview
Company
Impact
Profile match
Headquartered in New York City, United States, Vserve Ebusiness Solutions is a global provider of specialized e-commerce support and business process management services. The company offers a comprehensive range of solutions, including product catalog management, supply chain coordination, marketplace operations, and back-office financial support. Supported by offshore delivery centers across India and the Philippines, it helps enterprises and online retailers streamline workflows, improve data accuracy, and scale their digital operations efficiently.

Python Data Extraction Engineer - Web Scraping & Government Data

Location: Coimbatore

Experience: 3-7 years

Role Type: Full-time

About the Role

We are building a data intelligence platform that looks to analyze fragmented public information into structured, actionable business data.

We are looking for a strong Python Data Extraction Engineer who can discover, extract, clean, normalize and integrate data from government portals, public websites, PDFs, APIs and other open data sources.

This is not a conventional application-development role. The ideal candidate enjoys solving difficult data-acquisition problems involving poorly structured websites, inconsistent government portals, changing schemas, PDFs, JavaScript-rendered pages and large volumes of semi-structured information.

What You Will Own

You will build and maintain the data-acquisition layer of the platform.

Key responsibilities include:

Identify and evaluate government and public data sources relevant to property, businesses and commercial activity.

Build Python-based crawlers and extraction pipelines for government portals and public websites.

Extract structured information from HTML pages, tables, PDFs, downloadable files and publicly accessible APIs.

Work with JavaScript-rendered websites and multi-step public search interfaces.

Automate recurring extraction from multiple sources while respecting applicable access rules, rate limits, and terms.

Clean, standardize and normalize inconsistent data from different government authorities.

Perform entity matching and record linkage across datasets using fields such as owner name, company name, address, survey number, coordinates and project information.

Develop mechanisms to detect website/schema changes and extraction failures.

Build validation and QA processes to measure completeness and accuracy.

Store extracted information in structured databases and expose clean datasets to downstream applications.

Work closely with GIS, product and engineering teams to combine location-based signals with government/public records.

Research new public and open-data sources that can improve the accuracy and completeness of our intelligence.

Examples of Data Sources

The work may involve sources such as:

Municipal corporation portals

State planning and development authorities

DTCP and similar planning authorities

RERA databases

Building and planning permission records

Land and property records

Tender and procurement portals

Company/business registries

Government open-data portals

Environmental and regulatory approvals

Public notices and downloadable government documents

Maps and geospatial datasets

Other legally accessible public and open-source information

Required Technical Skills

Strong hands-on experience with:

Python

Web scraping and crawling

Requests / HTTP clients

BeautifulSoup / lxml

Selenium and/or Playwright

REST APIs and JSON

HTML/XML parsing

Pandas

SQL

Data cleaning and transformation

Regex and text processing

ETL/data pipelines

Git

What We Are Looking For

We particularly want someone who is a problem solver rather than simply a Python programmer.

The candidate should be able to investigate the available sources, understand how the underlying website works, determine the best extraction approach, build the pipeline and validate the resulting data.

Ideal Background

Candidates may come from backgrounds such as:

Web scraping / data extraction companies

Alternative-data companies

PropTech / real-estate data companies

Market-intelligence companies

OSINT/data intelligence companies

Government-data projects

Data aggregation platforms

Lead/data enrichment companies

GIS/location-intelligence companies

Success in This Role

Within the first few months, the successful candidate should be able to:

Map relevant government/public data sources.

Build reliable extraction pipelines across multiple portals.

Convert fragmented information into standardized records.

Cross-reference records from multiple sources.

Establish automated QA and monitoring.

Continuously discover additional datasets that improve our product's coverage and accuracy.

The objective is not simply to scrape websites. It is to build a scalable public-data acquisition and enrichment engine that becomes a core component of our intelligence platform.

Requirements

Python, pandas, webscraping, web crawling,selenium postman

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,229,538 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
Coimbatore
≈ $17k – $44k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Bengaluru
Python
JavaScript
Node JS
Frontend
React.js
DevOps
Rest API
Terraform
CloudFormation
CI/CD
AWS
Incident Management
IoT
MQTT
Management
Agile
Apply
≈ $16k – $42k per year (Estimated) • In office • 7+ years exp • Bengaluru
JavaScript
Java
SQL
PowerShell
Java
Hibernate
Databases
MySQL
DevOps
Rest API
CI/CD
IAM
SOAP
Cybersecurity
Active Directory
LDAP
Management
Agile
Apply
≈ $16k – $40k per year (Estimated) • In office • Full-Time • 3+ years exp • India
JavaScript
Java
TypeScript
Java
Spring Boot
Hibernate
Frontend
Angular
DevOps
Rest API
Git
Apply
≈ $16k – $40k per year (Estimated) • In office • Full-Time • 3+ years exp • India
JavaScript
Java
TypeScript
Java
Spring Boot
Hibernate
Frontend
Angular
DevOps
Rest API
CI/CD
Git
Apply
≈ $16k – $41k per year (Estimated) • In office • Full-Time • 5+ years exp • India
Java
Java
Spring Framework
Maven
Spring Boot
Gradle
Databases
Cassandra
DynamoDB
Couchbase
Apache Kafka
AI/ML
Claude
Prompt Engineering
Gemini
LLM
RAG
Semantic Search
OpenAI
Anthropic
GPT-4
Semantic Search
DevOps
Rest API
CI/CD
Jenkins
Git
Docker
Kubernetes
Apply
Remote (likely United States) • Full-Time • Bachelor's Degree
Python
JavaScript
Java
TypeScript
Node JS
Python
Django
Java
Spring Boot
Node JS
Express
Frontend
Vue.js
Angular
DevOps
GCP
Azure
CI/CD
AWS
Unix
Management
Agile
Apply
≈ $14k – $36k per year (Estimated) • In office • Full-Time • 1+ year exp • Bengaluru
Java
SQL
DevOps
Rest API
QA
Postman
Apply
$80k – $95k per year • In office • 5+ years exp • Bachelor's Degree • Boston
Python
SQL
R
Stata
Python
Flask
R
Shiny
AI/ML
Streamlit
Analytics
Power BI
Management
Microsoft Office
Apply
≈ $38k – $76k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Singapore
Python
SQL
Apply
AWS Developer 1 day ago
≈ $58k – $160k per year (Estimated) • Remote (location not specified) • Full-Time • 3+ years exp
Python
JavaScript
Node JS
Python
Boto3
DevOps
Terraform
CloudFormation
GitLab CI
CI/CD
Jenkins
AWS
AWS Lambda
Amazon S3
IAM
Amazon CloudWatch
Apply
Accounts Payable 1 day ago
In office • Full-Time • Bachelor's Degree • Coimbatore
AI/ML
OCR
DevOps
SLI/SLO/SLA
Analytics
Microsoft Excel
Management
Outlook
Apply
Graphics Designer 1 day ago
≈ $12k – $28k per year (Estimated) • In office • Full-Time • 3+ years exp • Coimbatore
Design
Adobe Photoshop
Figma
Canva
Apply
≈ $8k – $18k per year (Estimated) • Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • Coimbatore
SpaceTech
QGIS
Apply
In office • Full-Time • Bachelor's Degree • Coimbatore
Analytics
Microsoft Excel
Apply
≈ $11k – $30k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Coimbatore
Cybersecurity
Microsoft Defender
Microsoft Entra ID
SIEM
Apply
≈ $15k – $40k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Coimbatore
JavaScript
PHP
SQL
PHP
Laravel
Symfony
WordPress
WooCommerce
Magento
Databases
MySQL
Frontend
JQuery
Mobile
MVC
DevOps
Git
Bitbucket
SOAP
Apply
In office • Full-Time • Coimbatore
C#
C#
.NET
DevOps
Rest API
CI/CD
Git
Apply
In office • Full-Time • Coimbatore
JavaScript
TypeScript
SQL
C#
C#
ASP.NET Core
Blazor
Frontend
Vue.js
React.js
WebAssembly
Mobile
Dependency Injection
DevOps
Rest API
CI/CD
Git
Apply
In office • Full-Time • Bachelor's Degree • Coimbatore
MATLAB
Apply
≈ $20k – $49k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Coimbatore
C#
C#
.NET
AI/ML
AI Agents
DevOps
GCP
Azure DevOps
Azure
AWS
SLI/SLO/SLA
Analytics
Tableau
Power BI
Management
Slack
Confluence
SharePoint
Agile
Scrum
Waterfall
Apply
See all jobs
This is one of many
1,229,538 more open roles from verified company boards, updated every day.