D4Vinci/Scrapling
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!
Explore 9 GitHub repositories focused on scraping. Discover top-starred projects and those trending this week.
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!
🕵️♂️ Collect a dossier on a person by username from 3000+ sites
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
Twitter API Scraper | Without an API key | Twitter Internal API | Free | Twitter scraper | Twitter Bot
Linkedin Automation Tool: Describe your product. Define your target market. The AI finds the leads for you.
Learn step-by-step how to scrape Google Trends data and make a result comparison using Python and Oxylabs SERP API. Extract keywords, their popularity, breakdown by region, related queries, and more.
Low latency web data collector
A powerful Model Context Protocol (MCP) server that provides an all-in-one solution for public web access.
MinerU-HTML: An SLM-powered HTML main content extractor that outputs clean HTML bodies. Perfect for Deep Research Agents, RAG applications, and training data generation.
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!
🕵️♂️ Collect a dossier on a person by username from 3000+ sites
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
Twitter API Scraper | Without an API key | Twitter Internal API | Free | Twitter scraper | Twitter Bot
Linkedin Automation Tool: Describe your product. Define your target market. The AI finds the leads for you.
Learn step-by-step how to scrape Google Trends data and make a result comparison using Python and Oxylabs SERP API. Extract keywords, their popularity, breakdown by region, related queries, and more.
Low latency web data collector
A powerful Model Context Protocol (MCP) server that provides an all-in-one solution for public web access.
MinerU-HTML: An SLM-powered HTML main content extractor that outputs clean HTML bodies. Perfect for Deep Research Agents, RAG applications, and training data generation.