Capability

Web Scraping

11 classified repositories.

scrapy

scrapy

82Health
Source fact

Scrapy, a fast high-level web crawling & scraping framework for Python.

β˜… 64KPythonBSD-3-Clauseweb-scraping

firecrawl

firecrawl

79Health
Source fact

The context API to search, scrape, and interact with the web at scale. πŸ”₯

β˜… 171.6KTypeScriptAGPL-3.0web-scraping

apify

crawlee

78Health
Source fact

Crawleeβ€”A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.

β˜… 25.5KTypeScriptApache-2.0web-scraping

D4Vinci

Scrapling

77Health
Source fact

πŸ•·οΈ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!

β˜… 76.2KPythonBSD-3-Clauseweb-scraping

getmaxun

maxun

77Health
Source fact

πŸ”₯ The open-source no-code platform for web scraping, crawling, search and AI data extraction β€’ Turn websites into structured APIs in minutes πŸ”₯

β˜… 17.3KTypeScriptAGPL-3.0web-scraping

MontFerret

ferret

75Health
Source fact

Declarative data automation language and Go runtime for structured extraction workflows.

β˜… 6KGoApache-2.0web-scraping
75Health
Source fact

Crawleeβ€”A web scraping and browser automation library for Python to build reliable crawlers. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Parsel, BeautifulSoup, Playwright, and raw HTTP. Both headful and headless mode. With proxy rotation.

β˜… 9.5KPythonApache-2.0web-scraping
72Health
Source fact

πŸ”₯ Official Firecrawl MCP Server - Adds powerful web scraping and search to Cursor, Claude and any other LLM clients.

β˜… 7.3KJavaScriptMITweb-scraping

alirezamika

autoscraper

71Health
Source fact

A Smart, Automatic, Fast and Lightweight Web Scraper for Python

β˜… 7.9KPythonMITweb-scraping

code4craft

webmagic

64Health
Source fact

A scalable web crawler framework for Java.

β˜… 11.7KJavaApache-2.0web-scraping

brightdata

cli

60Health
Source fact

Official Bright Data CLI - scrape, search, and extract structured web data directly from your terminal.

β˜… 6.4KTypeScriptMITweb-scraping