全部标签
目录标签

#web-scraper

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

17 条记录
标签为“web-scraper”按健康指数排序
Packagist · PyPI
96卓越健康指数
php-curl-class/php-curl-class
PHP Curl Class makes it easy to send HTTP requests and integrate with web APIs
PHP★ 3,301↓ 140.4K/月2026年7月20日
Unlicense2026年7月20日 · 指标 2.10.0
PyPI
94卓越健康指数
D4Vinci/Scrapling
🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!
Python★ 72.6K2026年8月4日
BSD-3-Clause2026年8月4日 · 指标 2.10.0
PyPI
94卓越健康指数
ScrapeGraphAI/Scrapegraph-ai
Python scraper based on AI
Python★ 29K2026年8月5日
MIT2026年8月5日 · 指标 2.10.0
Packagist · Hex · crates.io +2
94卓越健康指数
firecrawl/firecrawl
The context API to search, scrape, and interact with the web at scale. 🔥
TypeScript · Python★ 161K↓ 3,278/月2026年8月4日
AGPL-3.02026年8月4日 · 指标 2.10.0
crates.io
87优秀健康指数
0x676e67/wreq
An ergonomic, privacy-aware Rust HTTP Client
Rust★ 1,004↓ 220.3K/月2026年8月28日
Apache-2.02026年8月28日 · 指标 2.10.0
npm · PyPI · crates.io
87优秀健康指数
us/crw
Fast, lightweight Firecrawl/Tavily alternative in Rust. Web scraper, crawler & search API with MCP server for AI agents. Drop-in Firecrawl-compatible API (/scrape, /crawl, /search). 2.3x faster than Tavily, 1.5x faster than Firecrawl in 1K-URL benchmarks. 6 MB RAM, single binary. Self-host or use managed cloud.
Rust★ 518↓ 9,070/月2026年8月3日
AGPL-3.02026年8月3日 · 指标 2.10.0
npm
86优秀健康指数
getmaxun/maxun
🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥
TypeScript★ 17K↓ 952/月2026年8月5日
AGPL-3.02026年8月5日 · 指标 2.10.0
Go
86优秀健康指数
gosom/google-maps-scraper
scrape data from Google Maps. Extracts data such as the name, address, phone number, website URL, rating, reviews number, latitude and longitude, reviews,email and more for each place
Go · HTML★ 5,6412026年8月28日
MIT2026年8月28日 · 指标 2.10.0
crates.io
78良好健康指数
0x676e67/wreq-util
Common utilities for wreq
Rust★ 90↓ 209.7K/月2026年9月5日
Apache-2.02026年9月5日 · 指标 2.10.0
npm
78良好健康指数
MrAdex77/google-play-scraper
Google Play scraper for Node.js with a fully typed TypeScript API. App details, search, top charts, reviews, permissions and data safety. A modern replacement for the unmaintained google-play-scraper.
TypeScript★ 5↓ 4,236/月2026年7月23日
MIT2026年7月23日 · 指标 2.10.0
npm · crates.io
75良好健康指数
0xMassi/webclaw
Fast, local-first web content extraction for LLMs. Scrape, crawl, extract structured data — all from Rust. CLI, REST API, and MCP server.
Rust★ 2,066↓ 424/月2026年7月26日
AGPL-3.02026年7月26日 · 指标 2.10.0
npm
71良好健康指数
stacksjs/ts-web-scraper
A powerful, type-safe web scraping library for TypeScript.
TypeScript★ 11↓ 177.5K/月2026年7月21日
MIT2026年7月21日 · 指标 2.10.0
Go
63中等健康指数
gan-of-culture/get-sauce
A command line program to download Hentai videos and images from multiple websites
Go★ 1952026年7月21日
MIT2026年7月21日 · 指标 2.10.0
npm · crates.io · Go +1
63中等健康指数
spider-rs/spider-clients
Python, Javascript, and Rust libraries for the Spider Cloud API.
Rust · Python · Go★ 26↓ 16.3K/月2026年7月16日
MIT2026年7月16日 · 指标 2.10.0
Go
48薄弱健康指数
leqwin/monloader
Online media downloader for monbooru
Go★ 32026年7月17日
AGPL-3.02026年7月17日 · 指标 2.10.0
PyPI
34存在风险健康指数
dipu-bd/lightnovel-crawler
Generate and download e-books from online sources.
Python★ 2,4942026年7月18日
GPL-3.02026年7月18日 · 指标 2.10.0
19危急健康指数
oxylabs/ai-scraper-py
AI Scraper is a powerful scraping tool and scrape agent built to automate data extraction with unmatched precision. Ideal for scalable AI scraping tasks across diverse web sources, this tool simplifies complex scraping operations into efficient, intelligent workflows.
混合★ 8502026年7月21日
无许可证2026年7月21日 · 指标 2.10.0