crawler

23 projetos partilham este topic do GitHub

crawler — firecrawl ★174.8kcrawlerScrapling — ★77.6kEasySpider — ★44.2kJobs_Applier_AI_Agent_AIHawk — ★30.3kScrapegraph-ai — ★30.2kcrawlee — ★25.6kmaxun — ★17.3kcrawlee-python — ★9.5kxiaobei — ★8.3kautoscraper — ★7.6ktrafilatura — ★6.7kmyGPTReader — ★4.4kscylla — ★4kAutoCrawler — ★1.7kfess — ★1.1kchatWeb — ★916crw — ★894hacker-news-digest — ★755Craw4LLM — ★664reader — ★558Fast-Powerful-Whisper-AI-Services-API — ★470scraperai — ★421extractor — ★320Scrapling★ 77.6kEasySpider★ 44.2kJobs_Applier_AI_Agent_AI…★ 30.3kScrapegraph-ai★ 30.2kcrawlee★ 25.6kmaxun★ 17.3kcrawlee-python★ 9.5kxiaobei★ 8.3kautoscraper★ 7.6ktrafilatura★ 6.7kmyGPTReader★ 4.4kscylla★ 4kAutoCrawler★ 1.7kfess★ 1.1kchatWeb★ 916crw★ 894hacker-news-digest★ 755Craw4LLM★ 664reader★ 558Fast-Powerful-Whisper-AI…★ 470scraperai★ 421extractor★ 320

Linhas conectam membros que estão mensuravelmente relacionados entre si. O tamanho do ponto reflete estrelas.

🧬 Membros
firecrawl
The context API to search, scrape, and interact with the web at scale. 🔥
★ 174.8k
Scrapling
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale…
★ 77.6k
EasySpider
A visual no-code/code-free web crawler/spider易采集:一个可视化浏览器自动化测试/数据采集/…
★ 44.2k
Jobs_Applier_AI_Agent_AIHawk
Open source AI job application bot in Python: browser automation and web scraping to read job postings, then…
★ 30.3k
Scrapegraph-ai
Python scraper based on AI
★ 30.2k
crawlee
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript…
★ 25.6k
maxun
🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn…
★ 17.3k
crawlee-python
Crawlee—A web scraping and browser automation library for Python to build reliable crawlers. Extract data…
★ 9.5k
xiaobei
为OPC/中小微企业量身打造的自媒体获客AI Agent
★ 8.3k
autoscraper
A Smart, Automatic, Fast and Lightweight Web Scraper for Python
★ 7.6k
trafilatura
Python & Command-line tool to gather text and metadata on the Web: Crawling, scraping, extraction, output as…
★ 6.7k
myGPTReader
A community-driven way to read and chat with AI bots - powered by chatGPT.
★ 4.4k
scylla
Intelligent proxy pool for Humans™ to extract content from the internet and build your own Large Language…
★ 4k
AutoCrawler
Google, Naver multiprocess image web crawler (Selenium)
★ 1.7k
fess
Open-source, self-hosted enterprise & site search server built on OpenSearch. Crawls web / file / DB / cloud…
★ 1.1k
chatWeb
ChatWeb can crawl web pages, read PDF, DOCX, TXT, and extract the main content, then answer your questions…
★ 916
crw
Fast, lightweight Firecrawl/Tavily alternative in Rust. Web scraper, crawler & search API with MCP server for…
★ 894
hacker-news-digest
:newspaper: Let ChatGPT Summarize Hacker News for You
★ 755
Craw4LLM
Official repository for "Craw4LLM: Efficient Web Crawling for LLM Pretraining"
★ 664
reader
Open source web infrastructure for AI. Scrape, crawl, and automate the web, clean markdown, browser sessions,…
★ 558
Fast-Powerful-Whisper-AI-Services-API
⚡ 一款用于自动语音识别 (ASR)、翻译的高性能异步 API。不需要购买Whisper…
★ 470
scraperai
ScraperAI is an open-source, AI-powered tool designed to simplify web scraping for users of all skill levels.
★ 421
extractor
Use LLMs to robustly extract web data
★ 320
🔗 Familias relacionadas

Medido a partir dos tópicos do GitHub compartilhados por ambos os projetos, ponderado pela raridade de cada tópico.