ocr

87 projetos partilham este topic do GitHub

ocr — PaddleOCR ★86.5kocrMinerU — ★76.2ktesseract — ★75.6kpaperless-ngx — ★43.7ktesseract.js — ★38.6kEasyOCR — ★29.8kopendataloader-pdf — ★28kLaTeX-OCR — ★16.5kunstructured — ★15.2kEasydict — ★14kbisheng — ★11.8kchandra — ★11.8kmanga-image-translator — ★10.2kX-AnyLabeling — ★9.9kBob — ★9.7kvideo-subtitle-extractor — ★9.3kailab — ★7.9kunstract — ★7kdonut — ★6.9kPaddleX — ★6.2kdoctr — ★6.2kSwift-AI — ★6klayout-parser — ★5.8kBallonsTranslator — ★5ksnow-shot — ★4.8kmmocr — ★4.7kTNN — ★4.6kjabref — ★4.5kRuVector — ★4.4kLLPlayer — ★4kdeep-text-recognition-benchmark — ★3.9kTTime — ★3.3ktext-extract-api — ★3.2kEAST — ★3.1kllm_aided_ocr — ★2.9kopenrecall — ★2.9kmarkdownify-mcp — ★2.9kcomic-translate — ★2.9kHands-On-AI-Engineering — ★2.8kcaptcha_break — ★2.8kmanga-ocr — ★2.7kMinerU★ 76.2ktesseract★ 75.6kpaperless-ngx★ 43.7ktesseract.js★ 38.6kEasyOCR★ 29.8kopendataloader-pdf★ 28kLaTeX-OCR★ 16.5kunstructured★ 15.2kEasydict★ 14kbisheng★ 11.8kchandra★ 11.8kmanga-image-translator★ 10.2kX-AnyLabeling★ 9.9kBob★ 9.7kvideo-subtitle-extractor★ 9.3kailab★ 7.9kunstract★ 7kdonut★ 6.9kPaddleX★ 6.2kdoctr★ 6.2kSwift-AI★ 6klayout-parser★ 5.8kBallonsTranslator★ 5ksnow-shot★ 4.8kmmocr★ 4.7kTNN★ 4.6kjabref★ 4.5kRuVector★ 4.4kLLPlayer★ 4kdeep-text-recognition-be…★ 3.9kTTime★ 3.3ktext-extract-api★ 3.2kEAST★ 3.1kllm_aided_ocr★ 2.9kopenrecall★ 2.9kmarkdownify-mcp★ 2.9kcomic-translate★ 2.9kHands-On-AI-Engineering★ 2.8kcaptcha_break★ 2.8kmanga-ocr★ 2.7k

Linhas conectam membros que estão mensuravelmente relacionados entre si. O tamanho do ponto reflete estrelas.

🧬 Membros
PaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that…
★ 86.5k
MinerU
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic…
★ 76.2k
tesseract
Tesseract Open Source OCR Engine (main repository)
★ 75.6k
paperless-ngx
A community-supported supercharged document management system: scan, index and archive all your documents
★ 43.7k
tesseract.js
Pure Javascript OCR for more than 100 Languages 📖🎉🖥
★ 38.6k
EasyOCR
Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese,…
★ 29.8k
opendataloader-pdf
PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
★ 28k
LaTeX-OCR
pix2tex: Using a ViT to convert images of equations into LaTeX code.
★ 16.5k
unstructured
Convert documents to structured data effortlessly. Unstructured is open-source ETL solution for transforming…
★ 15.2k
Easydict
一个简洁优雅的词典翻译 macOS App。开箱即用,支持离线 OCR…
★ 14k
bisheng
BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and…
★ 11.8k
chandra
OCR model that handles complex tables, forms, handwriting with full layout.
★ 11.8k
manga-image-translator
Translate manga/image 一键翻译各类图片内文字 https://cotrans.touhou.ai/ (no longer working)
★ 10.2k
X-AnyLabeling
Open-source AI-assisted annotation platform for images, videos, text, and multimodal data.
★ 9.9k
Bob
Bob 是一款 macOS 平台的翻译和 OCR 软件。
★ 9.7k
video-subtitle-extractor
★ 9.3k
ailab
Experience, Learn and Code the latest breakthrough innovations with Microsoft AI
★ 7.9k
unstract
LLM-Driven Extraction of Unstructured Data — Built for API Deployments & ETL Pipeline Workflows
★ 7k
donut
Official Implementation of OCR-free Document Understanding Transformer (Donut) and Synthetic Document…
★ 6.9k
PaddleX
All-in-One Development Tool based on PaddlePaddle
★ 6.2k
doctr
docTR (Document Text Recognition) - a seamless, high-performing & accessible library for OCR-related tasks…
★ 6.2k
Swift-AI
The Swift machine learning library.
★ 6k
layout-parser
A Unified Toolkit for Deep Learning Based Document Image Analysis
★ 5.8k
BallonsTranslator
深度学习辅助漫画翻译工具, 支持一键机翻和简单的图像/文本编辑 | Yet another…
★ 5k
snow-shot
超好用的截图工具
★ 4.8k
mmocr
OpenMMLab Text Detection, Recognition and Understanding Toolbox
★ 4.7k
TNN
TNN: developed by Tencent Youtu Lab and Guangying Lab, a uniform deep learning inference framework for…
★ 4.6k
jabref
Desktop app for managing BibTeX and BibLaTeX (.bib) libraries
★ 4.5k
RuVector
RuVector is a High Performance, Real-Time, Self-Learning Ai, Vector GNN, Memory DB built in Rust.
★ 4.4k
LLPlayer
The media player for language learning, with dual subtitles, AI-generated subtitles, real-time translation,…
★ 4k
deep-text-recognition-benchmark
Text recognition (optical character recognition) with deep learning methods, ICCV 2019
★ 3.9k
TTime
★ 3.3k
text-extract-api
Document (PDF, Word, PPTX ...) extraction and parse API using state of the art modern OCRs + Ollama supported…
★ 3.2k
EAST
A tensorflow implementation of EAST text detector
★ 3.1k
llm_aided_ocr
Enhances Tesseract OCR output using LLMs (local or API) for error correction, smart chunking, and markdown…
★ 2.9k
openrecall
OpenRecall is a fully open-source, privacy-first alternative to proprietary solutions like Microsoft's…
★ 2.9k
markdownify-mcp
A Model Context Protocol server for converting almost anything to Markdown
★ 2.9k
comic-translate
AI comic and manga translator app/browser extension for automatically translating comics, manga, manhwa, BDs,…
★ 2.9k
Hands-On-AI-Engineering
A curated collection of practical AI projects implementing OCR systems, RAG, AI agents, and other AI use…
★ 2.8k
captcha_break
验证码识别
★ 2.8k
manga-ocr
Optical character recognition for Japanese text, with the main focus being Japanese manga
★ 2.7k
paperless-gpt
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
★ 2.6k
awesome-deep-text-detection-recognition
A curated list of resources for text detection/recognition (optical character recognition ) with deep…
★ 2.5k
manga-translator-ui
基于 manga-image-translator 的开源漫画翻译工具。支持日/韩/美漫自动翻译,内置…
★ 2.4k
comfyui_LLM_party
LLM Agent Framework in ComfyUI includes MCP sever, Omost,GPT-sovits, ChatTTS,GOT-OCR2.0, and FLUX prompt…
★ 2.3k
SimpleHTR
Handwritten Text Recognition (HTR) system implemented with TensorFlow.
★ 2.2k
deepseek-ocr.rs
Rust multi‑backend OCR/VLM engine (DeepSeek‑OCR-1/2, PaddleOCR‑VL, DotsOCR) with DSQ quantization and…
★ 2.2k
crnn
Convolutional Recurrent Neural Network (CRNN) for image-based sequence recognition.
★ 2.1k
docext
An on-premises, OCR-free unstructured data extraction, markdown conversion and benchmarking toolkit.…
★ 2k
CRNN_Chinese_Characters_Rec
(CRNN) Chinese Characters Recognition.
★ 1.9k
ocrs
Rust library and CLI tool for OCR (extracting text from images)
★ 1.9k
AdvancedLiterateMachinery
A collection of original, innovative ideas and algorithms towards Advanced Literate Machinery. This project…
★ 1.8k
extractous
Fast and efficient unstructured data extraction. Written in Rust with bindings for many languages.
★ 1.8k
react-native-executorch
Declarative way to run AI models in React Native on device, powered by ExecuTorch.
★ 1.7k
MixTeX-Latex-OCR
MixTeX multimodal LaTeX, ZhEn, and, Table OCR. It performs efficient CPU-based inference in a local offline…
★ 1.6k
ExtractThinker
ExtractThinker is a Document Intelligence library for LLMs, offering ORM-style interaction for flexible and…
★ 1.6k
yomitoku
YomiTokuはAIを活用した日本語文書解析エンジンを提供するPythonパッケージです。…
★ 1.6k
awesome-document-understanding
A curated list of resources for Document Understanding (DU) topic
★ 1.5k
docstrange
Extract and convert data from any document, images, pdfs, word doc, ppt or URL into multiple formats…
★ 1.5k
ai-hands-on
A group of notebooks and other files which can help you learn AI from scratch.
★ 1.3k
Texo
A minimalist SOTA LaTeX OCR model with only 20M parameters, running in browser. Full training pipeline…
★ 874
note-companion
Note Companion: AI assistant for Obsidian that goes beyond just a chat. (prev File Organizer 2000)
★ 854
pdf-reader-mcp
Give your AI agent eyes for PDFs — structured text, tables, OCR, visual evidence, and page-level citations…
★ 842
aster
Recognizing cropped text in natural images.
★ 743
parseq
Scene Text Recognition with Permuted Autoregressive Sequence Models (ECCV 2022)
★ 727
aster.pytorch
ASTER in Pytorch
★ 682
Versatile-OCR-Program
Multi-modal OCR pipeline optimized for ML training (text, figure, math, tables, diagrams)
★ 677
fast-plate-ocr
Lightweight & fast OCR models for license plate text recognition.
★ 644
MakeMoneyWithAI
A list of open-source AI projects you can use to generate income easily.
★ 643
BetterOCR
🔍 Better text detection by combining multiple OCR engines (EasyOCR, Tesseract, and Pororo) with 🧠 LLM.
★ 636
AI-Bank-Statement-Document-Automation-By-LLM-And-Personal-Finanical-Analysis-Prediction
AI Bank Statement Document Automation By LLM model and Personal Finanical Analysis
★ 600
TurboOCR
TurboOCR, >200 img/s OmnidocBench. TensorRT FP16, PP-OCRv6, HTTP + gRPC
★ 580
ReMind
Your Local Artificial Memory on your Device.
★ 513
OCRAutoScore
OCR自动化阅卷项目
★ 483
GLM-skills
Official skills for the GLM family of models.
★ 452
ArkhamMirror
Local-first AI-powered document intelligence platform for investigative journalism
★ 449
light-ocr
Fast, offline OCR for Node.js & C++. PP-OCRv6 with Core ML / WebGPU hardware acceleration — recognize text…
★ 442
DeepSeek-OCR-WebUI
🎨 Ready-to-use DeepSeek-OCR Web UI | Modern Interface | 7 Recognition Modes | Batch Processing | …
★ 437
awesome-ocr-resources
A collection of resources (including the papers and datasets) of OCR (Optical Character Recognition).
★ 432
overlay-translator
★ 424
ANPR
license plate detection and recognition
★ 397
hms-ml-demo
HMS ML Demo provides an example of integrating Huawei ML Kit service into applications. This example…
★ 372
rowfill
Open-source spreadsheets platform for deep research and document processing
★ 368
awesome-computer-vision-resources
a collection of computer vision projects&tools. 计算机视觉方向项目和工具集合。
★ 355
DocDiff
ACM Multimedia 2023: DocDiff: Document Enhancement via Residual Diffusion Models. Also contains 1597 red…
★ 350
watch-skill
Video understanding and self-verification for AI agents. Turn videos, streams, and agent screen recordings…
★ 241
mdflux
Turn any document into clean, AI-ready Markdown. Local-first desktop app: reads scanned PDFs, batches…
★ 189
🔗 Familias relacionadas

Medido a partir dos tópicos do GitHub compartilhados por ambos os projetos, ponderado pela raridade de cada tópico.