Alle Tags
Katalog-Tag

#crawling

Alle Repositories im öffentlichen Register, die dieses Tag tragen — aus ihren GitHub-Topics oder den von ihren Paket-Registries veröffentlichten Schlagwörtern. Die Gesundheit wird nach derselben versionierten Methodik gemessen wie im übrigen Register.

9 Einträge
Getaggt als „crawling“Geordnet nach Gesundheitsindex
PyPI · npm
84GutGesundheitsindex
apify/apify-client-python
Apify API client for Python—Programmatically run Actors, manage and stream data from storages (datasets, key-value stores, request queues), schedule and monitor runs, and access the full Apify platform API. Sync and async interfaces with automatic retries and pagination.
Python★ 94↓ 2.5M/Monat21. Juli 2026
Apache-2.021. Juli 2026 · Metriken 1.13.0
npm
84GutGesundheitsindex
apify/crawlee
Crawlee—A web scraping and browser automation library for Node.js to build reliable crawlers. In JavaScript and TypeScript. Extract data for AI, LLMs, RAG, or GPTs. Download HTML, PDF, JPG, PNG, and other files from websites. Works with Puppeteer, Playwright, Cheerio, JSDOM, and raw HTTP. Both headful and headless mode. With proxy rotation.
TypeScript · MDX★ 24.8K19. Juli 2026
Apache-2.019. Juli 2026 · Metriken 1.13.0
PyPI
81GutGesundheitsindex
d4vinci/scrapling
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!
Python★ 69.4K↓ 863.2K/Monat14. Juli 2026
BSD-3-Clause14. Juli 2026 · Metriken 1.13.0
76GutGesundheitsindex
hardkoded/puppeteer-sharp
Headless Chrome .NET API
C#★ 3.90017. Juli 2026
MIT17. Juli 2026 · Metriken 1.13.0
PyPI · npm
68MittelGesundheitsindex
ArchiveBox/abx-dl
⬇️ A simple all-in-one CLI tool to download EVERYTHING from a URL (like youtube-dl/yt-dlp, forum-dl, gallery-dl, simpler ArchiveBox). 🎭 Uses headless Chrome to get HTML, JS, CSS, images/video/audio/subtitles, PDFs, screenshots, article text, git repos, and more...
Python★ 130↓ 9.160/Monat19. Juli 2026
MIT19. Juli 2026 · Metriken 1.13.0
crates.io · npm · Packagist +1
64MittelGesundheitsindex
xberg-io/crawlberg
High-performance web crawling engine with bindings for 11 languages
Rust★ 140↓ 0/Monat14. Juli 2026
MIT14. Juli 2026 · Metriken 1.13.0
PyPI
63MittelGesundheitsindex
adbar/courlan
Clean, filter and sample URLs to optimize data collection – Python & command-line – Deduplication, spam, content and language filters
Python★ 17721. Juli 2026
Apache-2.021. Juli 2026 · Metriken 1.13.0
Packagist
46GefährdetGesundheitsindex
crawlbase/crawlbase-php
A lightweight, dependency free PHP class that acts as wrapper for Crawlbase API
PHP★ 16↓ 2.779/Monat15. Juli 2026
Apache-2.015. Juli 2026 · Metriken 1.13.0
npm
35GefährdetGesundheitsindex
crawlbase/crawlbase-node
Fast dependency free library for Crawlbase API
JavaScript★ 919. Juli 2026
Apache-2.019. Juli 2026 · Metriken 1.13.0