CrawlLama 🦙 is an local AI agent that answers questions via Ollama and integrates web- and RAG-based research.
-
Updated
Oct 1, 2026 - Python
CrawlLama 🦙 is an local AI agent that answers questions via Ollama and integrates web- and RAG-based research.
An Amazon web scraper extracts product data like prices, reviews, and ratings using tools like BeautifulSoup or Scrapy, aiding in market research while adhering to ethical and legal guidelines.
SocioBlend adalah tool Python untuk meningkatkan views TikTok secara instan. Menggunakan antarmuka CLI berbasis rich, manajemen cookie dinamis, dan pengiriman otomatis >1.000 views tiap 15 menit. Cocok untuk eksplorasi automasi dan pembelajaran scraping etis.
Ethical web scraping library for public price monitoring with automatic robots.txt compliance
Python script using Selenium to scrape public LinkedIn profile data after login, saving it to CSV. Use responsibly and note data limitations.
Python Project demonstrating ethical web scraping with BeautifulSoup, Scrapy, and Playwright
A production-grade, async Python CLI tool to discover and extract local business records globally by keyword and location. Built with anti-bot bypass frameworks, rotating profiles, and incremental state saving.
Advanced Web Scraper is a comprehensive, production‑level web scraping framework built with Python. It demonstrates modern software engineering practices, robust architecture, and advanced features suitable for real‑world data extraction and monitoring tasks.
Production-grade scraping pipeline: Scrapy + Selenium, dual PostgreSQL/MongoDB storage, ethical scraping middleware, 28 automated tests with CI
A Windows-friendly, ethically-minded web scraper for educational/religious content. Includes safe downloads, 403 bypass testing, optional HTML→PDF/DOCX conversion, and detailed reporting.
Python-based web scraping and data analysis toolkit with automated data collection, processing, and visualization capabilities. Demonstrates data engineering and automation skills.
A polite, security-conscious web crawler & scraper in Python. Respects robots.txt and rate limits, guards against SSRF, and exports structured data as JSON Lines.
To associate your repository with the ethical-scraping topic, visit your repo's landing page and select "manage topics."