All tags
Catalogue tag

#pdf-processing

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

4 records
Tagged “pdf-processing”Ranked by health index
PyPI
77Goodhealth index
borb-pdf/borb
borb is a library for reading, creating and manipulating PDF files in python.
Python★ 3,569Aug 13, 2026
Custom licenseAug 13, 2026 · metrics 2.10.0
npm
63Moderatehealth index
mysleekdesigns/crawlforge-mcp
28 MCP tools that give Claude, Cursor and any MCP client the live web — scrape, crawl, search, real Google rank, change tracking, document parsing, plus an autonomous agent that researches from a plain-English prompt with no URLs. Clean Markdown and schema-validated JSON, not raw HTML. Local-Ollama extraction by default. MIT, 1,000 free credits.
JavaScript★ 2↓ 5,179/moSep 5, 2026
MITSep 5, 2026 · metrics 2.10.0
Go
60Moderatehealth index
giraffesyo/pdf
Robust, zero-dependency PDF text extraction for Go, with positioned glyphs, reading- order reconstruction, and hardened parsing.
Go★ 0Sep 5, 2026
MITSep 5, 2026 · metrics 2.10.0
npm
53Moderatehealth index
febbyRG/pdf-decomposer
A TypeScript Node.js library to parse all PDF page content (text, images, annotations, etc.) into JSON format.
TypeScript★ 2↓ 2,074/moJul 31, 2026
Custom licenseJul 31, 2026 · metrics 2.10.0