I build web scrapers and job-market data pipelines that run themselves. Python & Node.js · Playwright / Cheerio / Crawlee · anti-bot handling · Apify actors. Based in India 🇮🇳 · open to freelance data/scraping work.
🚀 Featured: Bharat Jobs Engine
An open, monthly, aggregates-only dataset of the Indian tech job market, built from public listings and refreshed automatically by a GitHub Actions cron.
From the latest snapshot: only ~10% of Indian tech listings even disclose a salary — and SQL, Python & Linux top the skills list. See the full report →
- 🧮 Aggregates only — k-anonymity suppression, no raw listings, no PII
- 🔁 Self-refreshing monthly via GitHub Actions
- 📊 CSV / JSON / Parquet + auto-generated charts · CC BY 4.0
| Actor | What it does |
|---|---|
| Naukri Jobs Scraper | Job listings from Naukri.com by keyword & location |
| Internshala Scraper | Internships & fresher jobs from Internshala |
| App Store & Google Play Reviews Scraper | User reviews from both app stores, any country |
Python · Node.js · Playwright · Cheerio · Crawlee · pandas
Anti-bot / stealth scraping · residential proxies · data cleaning & aggregation
Apify actors · GitHub Actions automation · clean CSV/Parquet/JSON pipelines
I build custom scrapers, clean market datasets, and pipelines that maintain themselves. Need data from a site that fights back, or a dataset delivered on a schedule?
→ Open an issue on any repo with the hire label, or check out my Apify actors. Projects typically start around $300.
Every scraper behind the Bharat Jobs Engine dataset is mine — it's a live sample of the work.