llm-scraper
Turn any webpage into structured data using LLMs
Extracting structured data from web pages to a schema in TypeScript, with the option to generate a reusable Playwright script from it.
You want a hosted scraping service; it is a library you run and supply a model to.
About llm-scraper
LLM Scraper is a TypeScript library that allows you to extract structured data from any webpage using LLMs.
Using the generate function you can generate re-usable playwright script that scrapes the contents according to a schema.
As an open-source project, we welcome contributions from the community. If you are experiencing any bugs or want to add some improvements, please feel free to open an issue or pull request.
llm-scraper is an open-source project written primarily in TypeScript, with 6.9k stars on GitHub. It was last updated in September 2026.
npm i zod playwright llm-scraperllm-scraper vs. the alternatives
All 21 alternatives →| Record | Stars | Pricing | ||
|---|---|---|---|---|
| llm-scraperthis listing | 6.9k | TypeScript | MIT | Open source |
| Browser Use | 117k | Python | MIT | Open source |
| UI-TARS-desktop | 39k | TypeScript | Apache-2.0 | Open source |
| page-agent | 29k | TypeScript | MIT | Open source |
| Stagehand | 25k | — | MIT | Open source |
| skyvern | 23k | Python | AGPL-3.0 | Open source |
