PyPI94卓越健康指数D4Vinci/Scrapling🕷️ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!Python★ 72.6K2026年8月4日BSD-3-Clause2026年8月4日 · 指标 2.10.0
npm94卓越健康指数inikulin/parse5HTML parsing/serialization toolset for Node.js. WHATWG HTML Living Standard (aka HTML5)-compliant.TypeScript★ 3,918↓ 802M/月2026年8月4日MIT2026年8月4日 · 指标 2.10.0
Go89优秀健康指数PuerkitoBio/goqueryA little like that j-thing, only in Go.Go · Roff★ 15K2026年8月4日BSD-3-Clause2026年8月4日 · 指标 2.10.0
PyPI89优秀健康指数soxoj/socid-extractor⛏️ The extraction engine behind Maigret: turn any profile URL into a structured OSINT record across 150+ sitesPython★ 1,067↓ 111.2K/月2026年8月22日MIT2026年8月22日 · 指标 2.10.0
PyPI78良好健康指数adbar/htmldateFast and robust date extraction from web pages, with Python or on the command-linePython★ 155↓ 15M/月2026年8月27日Apache-2.02026年8月27日 · 指标 2.10.0
PyPI54中等健康指数miso-belica/jusTextHeuristic based boilerplate removal toolPython★ 824↓ 12.3M/月2026年8月27日BSD-2-Clause2026年8月27日 · 指标 2.10.0