All tags
Catalogue tag

#crawlers

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

3 records
Tagged “crawlers”Ranked by health index
npm
98Exceptionalhealth index
apify/crawlee
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
TypeScript · MDX★ 25.2K↓ 6M/moAug 5, 2026
Apache-2.0Aug 5, 2026 · metrics 2.10.0
npm
83Excellenthealth index
omrilotan/isbot
🤖/👨‍🦰 Detect bots/crawlers/spiders using the user agent string
TypeScript · JavaScript★ 1,155↓ 102M/moAug 4, 2026
UnlicenseAug 4, 2026 · metrics 2.10.0
Maven
78Goodhealth index
Norconex/crawler
Norconex Crawlers (or spiders) are flexible web and filesystem crawlers for collecting, parsing, and manipulating data from the web or filesystem to various data repositories such as search engines.
Java★ 204Sep 5, 2026
Apache-2.0Sep 5, 2026 · metrics 2.10.0