All tags
Catalogue tag

#web-scraping

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

105 records
Tagged “web-scraping”Ranked by health index
Go
48Weakhealth index
leqwin/monloader
Online media downloader for monbooru
Go★ 3Jul 17, 2026
AGPL-3.0Jul 17, 2026 · metrics 2.10.0
PyPI
48Weakhealth index
sarperavci/UAForge
Generate statistically accurate User Agents and Client Hints (Sec-CH-UA). Deterministic, data-driven browser identities based on real-world market share distributions.
Python★ 9↓ 628/moAug 1, 2026
No licenseAug 1, 2026 · metrics 2.10.0
npm
47Weakhealth index
NovadaLabs/novada-mcp
One MCP server for all web data — search, scrape, crawl, proxy, and AI research in a single npx install. Works with Claude, Cursor, and any MCP client.
TypeScript · HTML★ 2↓ 5,804/moJul 16, 2026
No licenseJul 16, 2026 · metrics 2.10.0
Go
47Weakhealth index
khanglvm/online
Online web discovery CLI for AI agents
Go★ 0Jul 17, 2026
MITJul 17, 2026 · metrics 2.10.0
Packagist
42Weakhealth index
firecrawl/firecrawl-php
No repository description published.
PHP★ 2↓ 2,586/moAug 19, 2026
No licenseAug 19, 2026 · metrics 2.10.0
PyPI · npm
39Weakhealth index
dondai44423/master-fetch
MCP server for web fetching with Cloudflare bypass, Trafilatura extraction, and smart routing. Free, self-hosted, no API keys.
Python · HTML★ 4Jul 30, 2026
MITJul 30, 2026 · metrics 2.10.0
35Weakhealth index
vaaya-ai/vaaya-mcp
Vaaya MCP server — pay-per-call agent superpowers: media & video generation, product demo videos, web search & scraping, deep/market research, GTM & sales enrichment, code sandboxes, browser automation, email, memory. No API keys.
Shell★ 30Jul 19, 2026
MITJul 19, 2026 · metrics 2.10.0
PyPI
34At Riskhealth index
scrapy/scrapy
Scrapy, a fast high-level web crawling & scraping framework for Python.
Python★ 63.8KAug 12, 2026
BSD-3-ClauseAug 12, 2026 · metrics 2.10.0
Go
31At Riskhealth index
jim-ww/shiraberu
Multi-engine search aggregator for the terminal — no API keys, no server. CLI + Go library.
Go★ 1Jul 31, 2026
AGPL-3.0Jul 31, 2026 · metrics 2.10.0
PyPI · npm
29At Riskhealth index
omkarcloud/botasaurus-driver
Super Fast, Super Anti-Detect, and Super Intuitive Web Driver
Python★ 98Jul 30, 2026
MITJul 30, 2026 · metrics 2.10.0
Go
28At Riskhealth index
Synoppy/synoppy-go
Official Go SDK for Synoppy — the web-data layer for AI agents. Read, crawl, map, extract, classify & enrich any website on one key. Standard library only.
Go★ 1Jul 28, 2026
MITJul 28, 2026 · metrics 2.10.0
Packagist
28At Riskhealth index
roach-php/core
The complete web scraping toolkit for PHP.
PHP★ 1,455↓ 12K/moJul 27, 2026
No licenseJul 27, 2026 · metrics 2.10.0
Go
28At Riskhealth index
willygoid/fofa-grabber
A powerful Golang tool to fetch and export FOFA.info search results via API with real-time saving, concurrent requests, and multiple output formats (CSV, JSON, TXT)
Go★ 0Jul 17, 2026
MITJul 17, 2026 · metrics 2.10.0
PyPI
19Criticalhealth index
hosssein193m-ai/allykit
A powerful Python module offering excellent capabilities in web scraping, security, and meeting the standard needs of programmers.
Python★ 1Jul 26, 2026
MITJul 26, 2026 · metrics 2.10.0
19Criticalhealth index
oxylabs/how-to-scrape-google-finance
Use Web Scraper API to extract data from Google Finance, including stock titles, pricing, and price changes in percentages.
Python★ 1,200Jul 21, 2026
No licenseJul 21, 2026 · metrics 2.10.0