全部标签
目录标签

#metadata-extraction

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

6 条记录
标签为“metadata-extraction”按健康指数排序
npm
77良好健康指数
photostructure/exiftool-vendored.js
Fast, cross-platform Node.js access to ExifTool
TypeScript★ 552↓ 514.7K/月2026年7月15日
MIT2026年7月15日 · 指标 1.13.0
Packagist · crates.io · npm +1
73良好健康指数
xberg-io/xberg
A polyglot document intelligence framework with a Rust core. Extract text, metadata, images, and structured information from PDFs, Office documents, images, and 97+ formats. Available for Rust, Python, Ruby, Java, Go, PHP, Elixir, C#, R, C, TypeScript (Node/Bun/Wasm/Deno)- or use via CLI, REST API, or MCP server.
Rust★ 8,675↓ 40/月2026年7月20日
MIT2026年7月20日 · 指标 1.13.0
Maven
72良好健康指数
schemacrawler/SchemaCrawler
Free database schema discovery and comprehension tool
Java · HTML★ 1,8182026年7月17日
自定义许可证2026年7月17日 · 指标 1.13.0
PyPI
65中等健康指数
adbar/htmldate
Fast and robust date extraction from web pages, with Python or on the command-line
Python★ 154↓ 13.4M/月2026年7月21日
Apache-2.02026年7月21日 · 指标 1.13.0
PyPI
63中等健康指数
russalo/file-observer
Deterministic file observation for pipelines — one read-only pass over a directory emits a reproducible JSON manifest of every file's type, metadata, structure, and provenance.
Python★ 22026年7月17日
自定义许可证2026年7月17日 · 指标 1.13.0
Go
58中等健康指数
wasylq/fss
fss or Full Studio Scraper - scrape all videos metadata of your favourite performers or studios
Go★ 5↓ 0/月2026年7月13日
GPL-3.02026年7月13日 · 指标 1.13.0