ocr

44 projetos partilham este topic do GitHub

ocr — PaddleOCR ★88.6kocrMinerU — ★78.9ktesseract — ★75.4ktesseract.js — ★38.5kunstructured — ★15.4kbisheng — ★11.9kchandra — ★11.6kvideo-subtitle-extractor — ★9.2kunstract — ★7.2kdoctr — ★6.2ksnow-shot — ★4.8kLLPlayer — ★4.1kmodlens — ★3.8kTTime — ★3.3ktext-extract-api — ★3.2kEAST — ★3.1kllm_aided_ocr — ★3kHands-On-AI-Engineering — ★3kpaperless-gpt — ★2.7kawesome-deep-text-detection-recognition — ★2.5kSimpleHTR — ★2.2kdeepseek-ocr.rs — ★2.2kcrnn — ★2.1kCRNN_Chinese_Characters_Rec — ★1.9kocrs — ★1.9kextractous — ★1.8kExtractThinker — ★1.6kyomitoku — ★1.5kdocstrange — ★1.5kTurboOCR — ★1kMakeMoneyWithAI — ★888note-companion — ★852aster — ★742parseq — ★727aster.pytorch — ★682AI-Bank-Statement-Document-Automation-By-LLM-And-Personal-Finanical-Analysis-Prediction — ★625GLM-skills — ★469mdflux — ★411watch-skill — ★324pdf-mcp — ★117MinerU-Skill — ★108MinerU★ 78.9ktesseract★ 75.4ktesseract.js★ 38.5kunstructured★ 15.4kbisheng★ 11.9kchandra★ 11.6kvideo-subtitle-extractor★ 9.2kunstract★ 7.2kdoctr★ 6.2ksnow-shot★ 4.8kLLPlayer★ 4.1kmodlens★ 3.8kTTime★ 3.3ktext-extract-api★ 3.2kEAST★ 3.1kllm_aided_ocr★ 3kHands-On-AI-Engineering★ 3kpaperless-gpt★ 2.7kawesome-deep-text-detect…★ 2.5kSimpleHTR★ 2.2kdeepseek-ocr.rs★ 2.2kcrnn★ 2.1kCRNN_Chinese_Characters_…★ 1.9kocrs★ 1.9kextractous★ 1.8kExtractThinker★ 1.6kyomitoku★ 1.5kdocstrange★ 1.5kTurboOCR★ 1kMakeMoneyWithAI★ 888note-companion★ 852aster★ 742parseq★ 727aster.pytorch★ 682AI-Bank-Statement-Docume…★ 625GLM-skills★ 469mdflux★ 411watch-skill★ 324pdf-mcp★ 117MinerU-Skill★ 108

Linhas conectam membros que estão mensuravelmente relacionados entre si. O tamanho do ponto reflete estrelas.

🧬 Membros
PaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that…
★ 88.6k
MinerU
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic…
★ 78.9k
tesseract
Tesseract Open Source OCR Engine (main repository)
★ 75.4k
tesseract.js
Pure Javascript OCR for more than 100 Languages 📖🎉🖥
★ 38.5k
unstructured
Convert documents to structured data effortlessly. Unstructured is open-source ETL solution for transforming…
★ 15.4k
bisheng
BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and…
★ 11.9k
chandra
OCR model that handles complex tables, forms, handwriting with full layout.
★ 11.6k
video-subtitle-extractor
★ 9.2k
unstract
LLM-Driven Extraction of Unstructured Data — Built for API Deployments & ETL Pipeline Workflows
★ 7.2k
doctr
docTR (Document Text Recognition) - a seamless, high-performing & accessible library for OCR-related tasks…
★ 6.2k
snow-shot
超好用的截图工具
★ 4.8k
LLPlayer
The media player for language learning, with dual subtitles, AI-generated subtitles, real-time translation,…
★ 4.1k
modlens
The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste…
★ 3.8k
TTime
★ 3.3k
text-extract-api
Document (PDF, Word, PPTX ...) extraction and parse API using state of the art modern OCRs + Ollama supported…
★ 3.2k
EAST
A tensorflow implementation of EAST text detector
★ 3.1k
llm_aided_ocr
Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown…
★ 3k
Hands-On-AI-Engineering
A curated collection of practical AI projects implementing OCR systems, RAG, AI agents, and other AI use…
★ 3k
paperless-gpt
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
★ 2.7k
awesome-deep-text-detection-recognition
A curated list of resources for text detection/recognition (optical character recognition ) with deep…
★ 2.5k
SimpleHTR
Handwritten Text Recognition (HTR) system implemented with TensorFlow.
★ 2.2k
deepseek-ocr.rs
Rust multi‑backend OCR/VLM engine (DeepSeek‑OCR-1/2, PaddleOCR‑VL, DotsOCR) with DSQ quantization and…
★ 2.2k
crnn
Convolutional Recurrent Neural Network (CRNN) for image-based sequence recognition.
★ 2.1k
CRNN_Chinese_Characters_Rec
(CRNN) Chinese Characters Recognition.
★ 1.9k
ocrs
Rust library and CLI tool for OCR (extracting text from images)
★ 1.9k
extractous
Fast and efficient unstructured data extraction. Written in Rust with bindings for many languages.
★ 1.8k
ExtractThinker
ExtractThinker is a Document Intelligence library for LLMs, offering ORM-style interaction for flexible and…
★ 1.6k
yomitoku
YomiTokuはAIを活用した日本語文書解析エンジンを提供するPythonパッケージです。…
★ 1.5k
docstrange
Extract and convert data from any document, images, pdfs, word doc, ppt or URL into multiple formats…
★ 1.5k
TurboOCR
TurboOCR, >200 img/s OmnidocBench. TensorRT FP16, PP-OCRv6, HTTP + gRPC
★ 1k
MakeMoneyWithAI
A list of open-source AI projects you can use to generate income easily.
★ 888
note-companion
Note Companion: AI assistant for Obsidian that goes beyond just a chat. (prev File Organizer 2000)
★ 852
aster
Recognizing cropped text in natural images.
★ 742
parseq
Scene Text Recognition with Permuted Autoregressive Sequence Models (ECCV 2022)
★ 727
aster.pytorch
ASTER in Pytorch
★ 682
AI-Bank-Statement-Document-Automation-By-LLM-And-Personal-Finanical-Analysis-Prediction
AI Bank Statement Document Automation By LLM model and Personal Finanical Analysis
★ 625
GLM-skills
Official skills for the GLM family of models.
★ 469
mdflux
Turn any document into clean, AI-ready Markdown. Local-first desktop app: reads scanned PDFs, batches…
★ 411
watch-skill
Video understanding and self-verification for AI agents. Turn videos, streams, and agent screen recordings…
★ 324
pdf-mcp
MCP server that lets Claude Code and other AI agents work through large PDFs, and whole folders of them,…
★ 117
MinerU-Skill
AI-Native document parser: PDF, Office & images → clean Markdown with LaTeX, tables & OCR. Zero-dependency…
★ 108
DocMeld
Lightweight Doc-to-agent-ready knowledge pipeline. Three-stage Bronze→Silver→Gold architecture extracts…
★ 100
opencode-senses
The vision plugin for OpenCode that truly understands images. Inspect, read, and reason about any screenshot…
★ 82
dsh-vision-complete
给 DeepSeek 补上「眼睛和耳朵」的多模态视觉插件:看图 / OCR / 物体检测 / 视频理解…
★ 42
🔗 Familias relacionadas

Medido a partir dos tópicos do GitHub compartilhados por ambos os projetos, ponderado pela raridade de cada tópico.