You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Adaptive Python web scraping toolkit + MCP server for AI agents. Self-healing selectors that survive site changes, TLS-fingerprint stealth to bypass anti-bot filters, CSS/XPath parsing, and 24 built-in scrapers, clean, structured, LLM-ready data from any URL.
An automated Snakemake and .bat-driven workflow that scrapes, cleans, and serializes Yandex and Pinterest imagery into a high-performance Parquet dataset.
A Node.js script to fetch image URLs for products using the Google Custom Search API. Reads data from a JSON file, retrieves the images, and saves the results with image URLs in a new file.