All tags
Catalogue tag

#pdf-text-extraction

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

2 records
Tagged “pdf-text-extraction”Ranked by health index
Go
60Moderatehealth index
giraffesyo/pdf
Robust, zero-dependency PDF text extraction for Go, with positioned glyphs, reading- order reconstruction, and hardened parsing.
Go★ 0Sep 5, 2026
MITSep 5, 2026 · metrics 2.10.0
npm
53Moderatehealth index
febbyRG/pdf-decomposer
A TypeScript Node.js library to parse all PDF page content (text, images, annotations, etc.) into JSON format.
TypeScript★ 2↓ 2,074/moJul 31, 2026
Custom licenseJul 31, 2026 · metrics 2.10.0