PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
Extract text, tables, and images from any PDF file for use in AI applications.
pip install -U opendataloader-pdf

Excerpts from the project README on GitHub. Copyright and licensing remain with the respective authors.
Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.
Get an email alert on its next release or when it starts trending — never miss the moment.
Free · no card · unsubscribe anytimePDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
opendataloader-pdf has 27.4k stars on GitHub. It has been forked 2.6k times. opendataloader-pdf is written mainly in Java. It has been in active development since 2025. opendataloader-pdf is available under the Apache-2.0 license. Its main topics are a11y, accessibility, ai, bounding-box.
PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
opendataloader-pdf is an open-source project. It is released under the Apache-2.0 license.
Yes. opendataloader-pdf is free and open source — you can use, modify and self-host it.
opendataloader-pdf is available under the Apache-2.0 license.
opendataloader-pdf is written mainly in Java.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://olud.ai/project/opendataloader-project-opendataloader-pdf.html)
Measured from GitHub topics shared by both projects, weighted by how rare each topic is.