scrapy
scrapy
Scrapy, a fast high-level web crawling & scraping framework for Python.
Capability
11 classified repositories.
scrapy
Scrapy, a fast high-level web crawling & scraping framework for Python.
firecrawl
The context API to search, scrape, and interact with the web at scale. π₯
apify
CrawleeβA web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
D4Vinci
π·οΈ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!
getmaxun
π₯ The open-source no-code platform for web scraping, crawling, search and AI data extraction β’ Turn websites into structured APIs in minutes π₯
MontFerret
Declarative data automation language and Go runtime for structured extraction workflows.
apify
CrawleeβA web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.
firecrawl
π₯ Official Firecrawl MCP Server - Adds powerful web scraping and search to Cursor, Claude and any other LLM clients.
alirezamika
A Smart, Automatic, Fast and Lightweight Web Scraper for Python
code4craft
A scalable web crawler framework for Java.
brightdata
Official Bright Data CLI - scrape, search, and extract structured web data directly from your terminal.