全部标签
目录标签

#scraper

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

82 条记录
标签为“scraper”按健康指数排序
PyPI
71良好健康指数
0xMH/pyfunda
Python API wrapper for Funda.nl, the Dutch real estate platform. Reverse-engineered mobile API client. No scraping, no Selenium, no CAPTCHA.
Python★ 1752026年7月26日
AGPL-3.02026年7月26日 · 指标 2.10.0
Go
71良好健康指数
PxyUp/fitter
New way for collect information from the API's/Websites
Go · HTML★ 1322026年8月28日
MIT2026年8月28日 · 指标 2.10.0
PyPI
71良好健康指数
jr200-labs/polars-hist-db
(jetstream | file) --> polars dataframe <--> mariadb (bitemporal)
Python★ 1↓ 6,870/月2026年7月16日
MIT2026年7月16日 · 指标 2.10.0
PyPI
71良好健康指数
sdaqo/anipy-cli
Little tool in python to watch and download anime from the terminal (the better way to watch anime). Also applicable as an API
Python★ 486↓ 23.5K/月2026年7月19日
GPL-3.02026年7月19日 · 指标 2.10.0
npm
71良好健康指数
stacksjs/ts-web-scraper
A powerful, type-safe web scraping library for TypeScript.
TypeScript★ 11↓ 177.5K/月2026年7月21日
MIT2026年7月21日 · 指标 2.10.0
Maven
67良好健康指数
fanyong920/jvppeteer
Java API For Chrome and Firefox
Java★ 8082026年7月17日
Apache-2.02026年7月17日 · 指标 2.10.0
Go
67良好健康指数
gosom/scrapemate
Golang Crawling and scraping framework
Go★ 2072026年8月2日
MIT2026年8月2日 · 指标 2.10.0
npm · Go
67良好健康指数
pinchtab/seaportal
SeaPortal — Fast, HTTP-first content extraction for AI agents. No browser required.
Go · HTML★ 5↓ 207/月2026年7月17日
MIT2026年7月17日 · 指标 2.10.0
npm
63中等健康指数
laurengarcia/url-metadata
npm module: Fetch a url & scrape the metadata from its HTML with Node.js or the browser.
JavaScript★ 172↓ 95.4K/月2026年9月6日
MIT2026年9月6日 · 指标 2.10.0
npm
63中等健康指数
mysleekdesigns/crawlforge-mcp
28 MCP tools that give Claude, Cursor and any MCP client the live web — scrape, crawl, search, real Google rank, change tracking, document parsing, plus an autonomous agent that researches from a plain-English prompt with no URLs. Clean Markdown and schema-validated JSON, not raw HTML. Local-Ollama extraction by default. MIT, 1,000 free credits.
JavaScript★ 2↓ 5,179/月2026年9月5日
MIT2026年9月5日 · 指标 2.10.0
npm
63中等健康指数
rocktimsaikia/meta-fetcher
Scrape metadata from a website URL
TypeScript · JavaScript★ 147↓ 3,866/月2026年8月3日
MIT2026年8月3日 · 指标 2.10.0
npm · crates.io · Go +1
63中等健康指数
spider-rs/spider-clients
Python, Javascript, and Rust libraries for the Spider Cloud API.
Rust · Python · Go★ 26↓ 16.3K/月2026年7月16日
MIT2026年7月16日 · 指标 2.10.0
PyPI
62中等健康指数
farfarfun/funflix
影视资源分享文本的结构化采集、解析与网盘链接校验 - 采集/抽取/校验分层幂等,可单独重跑
Python★ 0↓ 2,003/月2026年8月30日
MIT2026年8月30日 · 指标 2.10.0
npm
62中等健康指数
recipe-scrapers/recipe-scrapers
A TypeScript library for scraping recipe data from cooking websites
TypeScript★ 6↓ 3,693/月2026年8月2日
MIT2026年8月2日 · 指标 2.10.0
Go
62中等健康指数
tamnd/ccrawl-cli
A fast, friendly command line for Common Crawl: URL index search, WARC fetch, Parquet columnar queries, and dataset building.
Go★ 82026年8月9日
Apache-2.02026年8月9日 · 指标 2.10.0
Go
60中等健康指数
AlexGustafsson/systembolaget-api
A cross-platform solution for using Systembolaget's APIs. For up-to-date data see https://github.com/AlexGustafsson/systembolaget-api-data.
Go · JavaScript★ 282026年7月26日
自定义许可证2026年7月26日 · 指标 2.10.0
Go
60中等健康指数
Pradumnasaraf/scrapy
Scrapy is a CLI tool used to scrape data from various websites
Go★ 42026年7月29日
Apache-2.02026年7月29日 · 指标 2.10.0
PyPI
60中等健康指数
bpodlipnik/mirror-url
Python CLI/library for mirroring and syncing solar-physics mission data (SOHO, STEREO, PROBA-3, SolarSoft, and similar HTTP(S) data repositories), with caching, filtering, and parallel downloads
Python★ 1↓ 2,582/月2026年8月1日
MIT2026年8月1日 · 指标 2.10.0
Go
60中等健康指数
man90es/BDO-REST-API
Scraper for Black Desert Online community data with a built-in API server.
Go★ 232026年7月25日
MIT2026年7月25日 · 指标 2.10.0
Packagist
59中等健康指数
boatracevibeproject/scraper
BVP Scraper は、ボートレースの公式サイトから出走表、直前情報、オッズ、結果をスクレイピングするための PHP ライブラリです。
PHP★ 3↓ 9,665/月2026年7月18日
MIT2026年7月18日 · 指标 2.10.0
npm
59中等健康指数
brandonkramer/pi-scraper
Pi extension for fast page scraping, recursive crawling, URL/site mapping, brand extraction, content diffing, PDF text extraction, and deterministic vertical extraction.
TypeScript · HTML★ 7↓ 607/月2026年8月28日
MIT2026年8月28日 · 指标 2.10.0
npm
59中等健康指数
cherchyk/MCPBrowser
该仓库未发布描述。
JavaScript★ 8↓ 1,837/月2026年7月15日
MIT2026年7月15日 · 指标 2.10.0
PyPI
57中等健康指数
codelucas/newspaper
newspaper3k is a news, full-text, and article metadata extraction in Python 3. Advanced docs:
Python★ 15.1K↓ 788.5K/月2026年8月12日
MIT2026年8月12日 · 指标 2.10.0
Go
57中等健康指数
kwadkore/ws-scraper
[en.]ws-tcg.com scraper. Permanently forked from https://github.com/Akenaide/wsoffcli
HTML · Go★ 02026年7月23日
Apache-2.02026年7月23日 · 指标 2.10.0
PyPI
56中等健康指数
c4road/earningspy
Your bot's best friend.
Python★ 22026年8月22日
MIT2026年8月22日 · 指标 2.10.0
Go
53中等健康指数
Dhairya3391/kari
Kari just a tui for getting media from providers and playing it in your media player with nice to have features.
Go★ 462026年9月5日
MIT2026年9月5日 · 指标 2.10.0
Go
53中等健康指数
tamnd/martinfowler-cli
Fetch Martin Fowler's technical articles, bliki posts, and book entries from the terminal
Go · Python★ 02026年7月22日
Apache-2.02026年7月22日 · 指标 2.10.0
Go
53中等健康指数
tamnd/neuralnetworksdl-cli
Read Neural Networks and Deep Learning book chapters and pages as JSON from the command line
Go · Python★ 02026年7月21日
Apache-2.02026年7月21日 · 指标 2.10.0
Go
53中等健康指数
tamnd/thegradient-cli
Fetch The Gradient AI research publication posts and listings from the terminal
Go · Python★ 02026年7月21日
Apache-2.02026年7月21日 · 指标 2.10.0
Go
53中等健康指数
tamnd/theodinproject-cli
Browse The Odin Project learning paths and lessons as JSON from the command line
Go · Python★ 02026年7月21日
Apache-2.02026年7月21日 · 指标 2.10.0