Todas las etiquetas
Etiqueta del catálogo

#html-parser

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

15 registros
Con la etiqueta «html-parser»Ordenado por índice de salud
npm
94Excepcionalíndice de salud
inikulin/parse5
HTML parsing/serialization toolset for Node.js. WHATWG HTML Living Standard (aka HTML5)-compliant.
TypeScript★ 3918↓ 802M/mes4 ago 2026
MIT4 ago 2026 · métricas 2.10.0
npm
91Excelenteíndice de salud
fb55/htmlparser2
The fast & forgiving HTML and XML parser
TypeScript★ 4785↓ 346M/mes4 ago 2026
MIT4 ago 2026 · métricas 2.10.0
PyPI · crates.io · npm
81Excelenteíndice de salud
bug-ops/scrape-rs
🦀 High-performance HTML parsing library. Rust core with native bindings for Python, Node.js & WASM. SIMD-accelerated, memory-safe, consistent API everywhere.
Rust★ 10↓ 6257/mes5 sept 2026
Apache-2.05 sept 2026 · métricas 2.10.0
Hex
81Excelenteíndice de salud
philss/floki
Floki is a simple HTML parser that enables search for nodes using CSS selectors.
Elixir★ 2145↓ 423.3K/mes17 jul 2026
MIT17 jul 2026 · métricas 2.10.0
Packagist
71Buenoíndice de salud
voku/simple_html_dom
📜 Modern Simple HTML DOM Parser for PHP
PHP★ 900↓ 226.4K/mes15 jul 2026
MIT15 jul 2026 · métricas 2.10.0
crates.io · npm · PyPI
69Buenoíndice de salud
nabbisen/mdka-rs
A HTML to Markdown (MD) converter balances conversion quality with runtime efficiency.
Rust · JavaScript · Python★ 54↓ 5080/mes3 ago 2026
Apache-2.03 ago 2026 · métricas 2.10.0
Go
65Buenoíndice de salud
dotcommander/defuddle
Go library and CLI for extracting web page content — articles, metadata, and clean text from any URL
Go★ 217 jul 2026
MIT17 jul 2026 · métricas 2.10.0
npm
63Moderadoíndice de salud
mysleekdesigns/crawlforge-mcp
28 MCP tools that give Claude, Cursor and any MCP client the live web — scrape, crawl, search, real Google rank, change tracking, document parsing, plus an autonomous agent that researches from a plain-English prompt with no URLs. Clean Markdown and schema-validated JSON, not raw HTML. Local-Ollama extraction by default. MIT, 1,000 free credits.
JavaScript★ 2↓ 5179/mes5 sept 2026
MIT5 sept 2026 · métricas 2.10.0
NuGet
62Moderadoíndice de salud
zzzprojects/html-agility-pack
Html Agility Pack (HAP) is a free and open-source HTML parser written in C# to read/write DOM and supports plain XPATH or XSLT. It is a .NET code library that allows you to parse "out of the web" HTML files.
C# · HTML★ 284622 ago 2026
MIT22 ago 2026 · métricas 2.10.0
Hex
60Moderadoíndice de salud
rusterlium/html5ever_elixir
NIF wrapper of html5ever using Rustler
HTML · Rust · Elixir★ 87↓ 7019/mes17 jul 2026
Apache-2.017 jul 2026 · métricas 2.10.0
PyPI
54Moderadoíndice de salud
miso-belica/jusText
Heuristic based boilerplate removal tool
Python★ 824↓ 12.3M/mes27 ago 2026
BSD-2-Clause27 ago 2026 · métricas 2.10.0
Packagist · npm
50Moderadoíndice de salud
duzun/hQuery.php
An extremely fast web scraper that parses megabytes of invalid HTML in a blink of an eye. PHP5.3+, no dependencies.
PHP★ 360↓ 12.6K/mes29 jul 2026
MIT29 jul 2026 · métricas 2.10.0
NuGet
35Débilíndice de salud
SoftCircuits/HtmlMonkey
Lightweight HTML/XML parser written in C#.
C#★ 6231 jul 2026
Licencia propia31 jul 2026 · métricas 2.10.0
Go
11Críticoíndice de salud
andreychh/tgen
Turns the Telegram Bot API HTML documentation into ready-to-use API bindings.
Python · Go★ 220 jul 2026
MIT20 jul 2026 · métricas 2.10.0
Go
7Críticoíndice de salud
wmentor/html
HTML data fetcher
Go★ 13 sept 2026
MIT3 sept 2026 · métricas 2.10.0