Document (PDF, Word, PPTX ...) extraction and parse API using state of the art modern OCRs + Ollama supported models. Anonymize documents. Remove PII. Convert any document or picture to structured JSON or Markdown
Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.
Document (PDF, Word, PPTX ...) extraction and parse API using state of the art modern OCRs + Ollama supported models. Anonymize documents. Remove PII. Convert any document or picture to structured JSON or Markdown
Document (PDF, Word, PPTX ...) extraction and parse API using state of the art modern OCRs + Ollama supported models. Anonymize documents. Remove PII. Convert any document or picture to structured JSON or Markdown
text-extract-api is an open-source project. It is released under the MIT license.
Yes. text-extract-api is free and open source — you can use, modify and self-host it.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://opensourceai.tech/project/catchthetornado-text-extract-api.html)