Todas las etiquetas
Etiqueta del catálogo

#spider

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

17 registros
Con la etiqueta «spider»Ordenado por índice de salud
PyPI
94Excepcionalíndice de salud
akfamily/akshare
AKShare is an elegant and simple financial data interface library for Python, built for human beings! 开源财经数据接口库
Python · JavaScript★ 21.8K5 ago 2026
MIT5 ago 2026 · métricas 2.10.0
PyPI
87Excelenteíndice de salud
DedSecInside/TorBot
Dark Web OSINT Tool
Python★ 472628 ago 2026
Licencia propia28 ago 2026 · métricas 2.10.0
crates.io
87Excelenteíndice de salud
spider-rs/spider
Get web data for AI agents and LLMs
Rust★ 2669↓ 42K/mes22 ago 2026
MIT22 ago 2026 · métricas 2.10.0
Packagist
80Excelenteíndice de salud
JayBizzle/Crawler-Detect
🕷 CrawlerDetect is a PHP class for detecting bots/crawlers/spiders via the user agent
PHP★ 2401↓ 2.6M/mes27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
Maven
78Buenoíndice de salud
Norconex/crawler
Norconex Crawlers (or spiders) are flexible web and filesystem crawlers for collecting, parsing, and manipulating data from the web or filesystem to various data repositories such as search engines.
Java★ 2045 sept 2026
Apache-2.05 sept 2026 · métricas 2.10.0
Go
67Buenoíndice de salud
gosom/scrapemate
Golang Crawling and scraping framework
Go★ 2072 ago 2026
MIT2 ago 2026 · métricas 2.10.0
RubyGems
63Moderadoíndice de salud
loadkpi/crawler_detect
Ruby gem to detect bots and crawlers via the user agent
Ruby★ 1486 sept 2026
MIT6 sept 2026 · métricas 2.10.0
npm · crates.io · Go +1
63Moderadoíndice de salud
spider-rs/spider-clients
Python, Javascript, and Rust libraries for the Spider Cloud API.
Rust · Python · Go★ 26↓ 16.3K/mes16 jul 2026
MIT16 jul 2026 · métricas 2.10.0
NuGet
62Moderadoíndice de salud
zzzprojects/html-agility-pack
Html Agility Pack (HAP) is a free and open-source HTML parser written in C# to read/write DOM and supports plain XPATH or XSLT. It is a .NET code library that allows you to parse "out of the web" HTML files.
C# · HTML★ 284622 ago 2026
MIT22 ago 2026 · métricas 2.10.0
PyPI
60Moderadoíndice de salud
moskrc/crawlerdetect
🕷CrawlerDetect is a Python library designed to identify bots, crawlers, and spiders by analyzing their user agents.
Python★ 4429 jul 2026
MIT29 jul 2026 · métricas 2.10.0
Packagist
57Moderadoíndice de salud
JayBizzle/Laravel-Crawler-Detect
A Laravel wrapper for CrawlerDetect - the web crawler detection library
PHP★ 324↓ 61.2K/mes28 jul 2026
MIT28 jul 2026 · métricas 2.10.0
Go
56Moderadoíndice de salud
x-way/crawlerdetect
Golang module to detect bots and crawlers via the user agent
Go★ 7017 jul 2026
MIT17 jul 2026 · métricas 2.10.0
48Débilíndice de salud
Senparc/Senparc.CO2NET
Base Common Library, support for.NET Framework &.NET Core
C#★ 36518 jul 2026
Apache-2.018 jul 2026 · métricas 2.10.0
Packagist
41Débilíndice de salud
zorlan/skycaiji
蓝天采集器是一款开源免费的爬虫系统,仅需点选编辑规则即可采集数据,可运行在本地、虚拟主机或云服务器中,几乎能采集所有类型的网页,无缝对接各类CMS建站程序,免登录实时发布数据,全自动无需人工干预!是网页大数据采集软件中完全跨平台的云端爬虫系统
PHP★ 207827 jul 2026
Licencia propia27 jul 2026 · métricas 2.10.0
NuGet
35Débilíndice de salud
SoftCircuits/HtmlMonkey
Lightweight HTML/XML parser written in C#.
C#★ 6231 jul 2026
Licencia propia31 jul 2026 · métricas 2.10.0
Maven
23En riesgoíndice de salud
ssssssss-team/spider-flow
新一代爬虫平台,以图形化方式定义爬虫流程,不写代码即可完成爬虫。
Java★ 11.4K27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
PyPI
22En riesgoíndice de salud
0-EternalJunior-0/GraphCrawler
Python бібліотека для сканування веб-сайтів та побудови графу їх структури.
Python · HTML★ 1↓ 4246/mes1 ago 2026
MIT1 ago 2026 · métricas 2.10.0