全部标签
目录标签

#pdf-parsing

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

5 条记录
标签为“pdf-parsing”按健康指数排序
PyPI
98卓越健康指数
py-pdf/pypdf
A pure-python PDF library capable of splitting, merging, cropping, and transforming the pages of PDF files
Python★ 10.2K2026年8月28日
自定义许可证2026年8月28日 · 指标 2.10.0
npm · Maven · PyPI
95卓越健康指数
opendataloader-project/opendataloader-pdf
PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.
Java · Python★ 28.2K↓ 54.5K/月2026年8月5日
Apache-2.02026年8月5日 · 指标 2.10.0
crates.io · PyPI · npm +5
93卓越健康指数
yfedoseev/pdf_oxide
The fastest PDF library for Python and Rust. Text extraction, image extraction, markdown conversion, PDF creation & editing. 0.8ms mean, 5× faster than industry leaders, 100% pass rate on 3,830 PDFs. MIT/Apache-2.0.
Rust★ 1,015↓ 391.7K/月2026年9月5日
Apache-2.02026年9月5日 · 指标 2.10.0
npm
80优秀健康指数
LibPDF-js/core
A modern PDF library for TypeScript. Parse, modify, and generate PDFs with a clean, intuitive API.
TypeScript★ 1,771↓ 352.6K/月2026年7月28日
MIT2026年7月28日 · 指标 2.10.0
PyPI
77良好健康指数
jsvine/pdfplumber
Plumb a PDF for detailed information about each char, rectangle, line, et cetera — and easily extract text and tables.
Python★ 10.7K2026年8月28日
MIT2026年8月28日 · 指标 2.10.0