Home Projects extractor
extractor
TypeScript

extractor

Use LLMs to robustly extract web data

by lightfeed · GitHub
Stars
Forks
License
Created
Last commit
Category
Language
ai-agentsarticle-extractorcrawlerApache-2.0TypeScript
View on GitHub
In plain words

Extract structured data from web pages using natural language prompts with this library.

You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
extractor — GitHub preview card
📈 Star history
320319
2026-07-202026-08-31
📈 Track extractor

Get an email alert on its next release or when it starts trending — never miss the moment.

Free · no card · unsubscribe anytime
Get email alerts →
📄 About

Use LLMs to robustly extract web data

extractor has 320 stars on GitHub. It has been forked 10 times. extractor is written mainly in TypeScript. It has been in active development since 2025. extractor is available under the Apache-2.0 license. Its main topics are ai-agents, article-extractor, crawler, data-engineering.

Frequently asked questions

What is extractor?

Use LLMs to robustly extract web data

Is extractor open source?

extractor is an open-source project. It is released under the Apache-2.0 license.

Is extractor free?

Yes. extractor is free and open source — you can use, modify and self-host it.

What license does extractor use?

extractor is available under the Apache-2.0 license.

What language is extractor written in?

extractor is written mainly in TypeScript.

🏅 Maintainer of this project?
olud.ai badge — extractor

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![olud.ai](https://olud.ai/badge.php?tool=lightfeed-extractor)](https://olud.ai/project/lightfeed-extractor.html)
More badge options →
🧬 Shares DNA with🧬 View the DNA map →

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.