crawler

24 projets partagent ce topic GitHub

crawler — firecrawl ★157.9kcrawlerScrapling — ★71.7kEasySpider — ★44.3kScrapegraph-ai — ★28.8kcrawlee — ★25.1kmaxun — ★16.9kcrawlee-python — ★9.4kautoscraper — ★7.7ktrafilatura — ★6.4kmyGPTReader — ★4.4kscylla — ★4kLeaked-GPTs — ★2.5kwebclaw — ★2.1kx-crawl — ★1.9kAutoCrawler — ★1.7kfess — ★1.1kchatWeb — ★916hacker-news-digest — ★756Craw4LLM — ★660reader — ★552crw — ★495Fast-Powerful-Whisper-AI-Services-API — ★470scraperai — ★424extractor — ★319Scrapling★ 71.7kEasySpider★ 44.3kScrapegraph-ai★ 28.8kcrawlee★ 25.1kmaxun★ 16.9kcrawlee-python★ 9.4kautoscraper★ 7.7ktrafilatura★ 6.4kmyGPTReader★ 4.4kscylla★ 4kLeaked-GPTs★ 2.5kwebclaw★ 2.1kx-crawl★ 1.9kAutoCrawler★ 1.7kfess★ 1.1kchatWeb★ 916hacker-news-digest★ 756Craw4LLM★ 660reader★ 552crw★ 495Fast-Powerful-Whisper-AI…★ 470scraperai★ 424extractor★ 319

Les traits relient les membres réellement apparentés entre eux. La taille des points suit les étoiles.

🧬 Membres
firecrawl
The API to search, scrape, and interact with the web at scale. 🔥
★ 157.9k
Scrapling
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale…
★ 71.7k
EasySpider
A visual no-code/code-free web crawler/spider易采集:一个可视化浏览器自动化测试/数据采集/…
★ 44.3k
Scrapegraph-ai
Python scraper based on AI
★ 28.8k
crawlee
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript…
★ 25.1k
maxun
🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn…
★ 16.9k
crawlee-python
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data…
★ 9.4k
autoscraper
A Smart, Automatic, Fast and Lightweight Web Scraper for Python
★ 7.7k
trafilatura
Python & Command-line tool to gather text and metadata on the Web: Crawling, scraping, extraction, output as…
★ 6.4k
myGPTReader
A community-driven way to read and chat with AI bots - powered by chatGPT.
★ 4.4k
scylla
Intelligent proxy pool for Humans™ to extract content from the internet and build your own Large Language…
★ 4k
Leaked-GPTs
Leaked GPTs Prompts Bypass the 25 message limit or to try out GPTs without a Plus subscription.
★ 2.5k
webclaw
Fast, local-first web content extraction for LLMs. Scrape, crawl, extract structured data — all from Rust.…
★ 2.1k
x-crawl
Flexible Node.js AI-assisted crawler library
★ 1.9k
AutoCrawler
Google, Naver multiprocess image web crawler (Selenium)
★ 1.7k
fess
Open-source, self-hosted enterprise & site search server built on OpenSearch. Crawls web / file / DB / cloud…
★ 1.1k
chatWeb
ChatWeb can crawl web pages, read PDF, DOCX, TXT, and extract the main content, then answer your questions…
★ 916
hacker-news-digest
:newspaper: Let ChatGPT Summarize Hacker News for You
★ 756
Craw4LLM
Official repository for "Craw4LLM: Efficient Web Crawling for LLM Pretraining"
★ 660
reader
Open source web infrastructure for AI. Scrape, crawl, and automate the web, clean markdown, browser sessions,…
★ 552
crw
Fast, lightweight Firecrawl/Tavily alternative in Rust. Web scraper, crawler & search API with MCP server for…
★ 495
Fast-Powerful-Whisper-AI-Services-API
⚡ 一款用于自动语音识别 (ASR)、翻译的高性能异步 API。不需要购买Whisper…
★ 470
scraperai
ScraperAI is an open-source, AI-powered tool designed to simplify web scraping for users of all skill levels.
★ 424
extractor
Use LLMs to robustly extract web data
★ 319
🔗 Familles voisines

Mesuré à partir des topics GitHub communs aux deux projets, pondérés par leur rareté.