全部标签
目录标签

#table-extraction

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

4 条记录
标签为“table-extraction”按健康指数排序
PyPI
76良好健康指数
pymupdf/pymupdf
PyMuPDF is a high performance Python library for data extraction, analysis, conversion & manipulation of PDF (and other) documents.
Python · SWIG★ 10.2K2026年7月17日
AGPL-3.02026年7月17日 · 指标 1.13.0
Packagist · crates.io · npm +1
73良好健康指数
xberg-io/xberg
A polyglot document intelligence framework with a Rust core. Extract text, metadata, images, and structured information from PDFs, Office documents, images, and 97+ formats. Available for Rust, Python, Ruby, Java, Go, PHP, Elixir, C#, R, C, TypeScript (Node/Bun/Wasm/Deno)- or use via CLI, REST API, or MCP server.
Rust★ 8,675↓ 40/月2026年7月20日
MIT2026年7月20日 · 指标 1.13.0
crates.io
69中等健康指数
bzsanti/oxidizePdf
Pure Rust PDF library for AI/RAG: structure-aware chunking, no ML, no C deps.
Rust★ 182↓ 6,045/月2026年7月13日
MIT2026年7月13日 · 指标 1.13.0
PyPI
63中等健康指数
jsvine/pdfplumber
Plumb a PDF for detailed information about each char, rectangle, line, et cetera — and easily extract text and tables.
Python★ 10.6K2026年7月17日
MIT2026年7月17日 · 指标 1.13.0