全部标签
目录标签

#pdf-processing

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

4 条记录
标签为“pdf-processing”按健康指数排序
PyPI
77良好健康指数
borb-pdf/borb
borb is a library for reading, creating and manipulating PDF files in python.
Python★ 3,5692026年8月13日
自定义许可证2026年8月13日 · 指标 2.10.0
npm
63中等健康指数
mysleekdesigns/crawlforge-mcp
28 MCP tools that give Claude, Cursor and any MCP client the live web — scrape, crawl, search, real Google rank, change tracking, document parsing, plus an autonomous agent that researches from a plain-English prompt with no URLs. Clean Markdown and schema-validated JSON, not raw HTML. Local-Ollama extraction by default. MIT, 1,000 free credits.
JavaScript★ 2↓ 5,179/月2026年9月5日
MIT2026年9月5日 · 指标 2.10.0
Go
60中等健康指数
giraffesyo/pdf
Robust, zero-dependency PDF text extraction for Go, with positioned glyphs, reading- order reconstruction, and hardened parsing.
Go★ 02026年9月5日
MIT2026年9月5日 · 指标 2.10.0
npm
53中等健康指数
febbyRG/pdf-decomposer
A TypeScript Node.js library to parse all PDF page content (text, images, annotations, etc.) into JSON format.
TypeScript★ 2↓ 2,074/月2026年7月31日
自定义许可证2026年7月31日 · 指标 2.10.0