全部标签
目录标签

#scraping

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

82 条记录
标签为“scraping”按健康指数排序
NuGet
48薄弱健康指数
darrylwhitmore/NScrape
A web scraping framework for .NET
C#★ 662026年7月17日
MIT2026年7月17日 · 指标 2.10.0
PyPI
48薄弱健康指数
sarperavci/UAForge
Generate statistically accurate User Agents and Client Hints (Sec-CH-UA). Deterministic, data-driven browser identities based on real-world market share distributions.
Python★ 9↓ 628/月2026年8月1日
无许可证2026年8月1日 · 指标 2.10.0
Go
47薄弱健康指数
anatolykoptev/go-threads
Threads (threads.net) scraping module — public GraphQL API, LSD token auth, go-stealth TLS fingerprinting
Go★ 22026年8月13日
Apache-2.02026年8月13日 · 指标 2.10.0
PyPI
47薄弱健康指数
online-judge-tools/api-client
API client to develop tools for competitive programming
Python★ 822026年7月15日
MIT2026年7月15日 · 指标 2.10.0
PyPI
47薄弱健康指数
ultrafunkamsterdam/nodriver
Successor of Undetected-Chromedriver. Providing a blazing fast framework for web automation, webscraping, bots and any other creative ideas which are normally hindered by annoying anti bot systems like Captcha / CloudFlare / Imperva / hCaptcha
Python★ 4,6482026年8月12日
AGPL-3.02026年8月12日 · 指标 2.10.0
Packagist
42薄弱健康指数
crawlbase/crawlbase-php
A lightweight, dependency free PHP class that acts as wrapper for Crawlbase API
PHP★ 16↓ 2,779/月2026年7月15日
Apache-2.02026年7月15日 · 指标 2.10.0
PyPI
41薄弱健康指数
DaKheera47/scraperecon
CLI recon tool for scraper developers. Detects TLS fingerprinting, JS challenges, bot protection, and rate limits across 4 stages
Python★ 42↓ 39/月2026年8月1日
MIT2026年8月1日 · 指标 2.10.0
npm
39薄弱健康指数
monid-ai/cli
该仓库未发布描述。
TypeScript★ 1↓ 2,563/月2026年7月19日
无许可证2026年7月19日 · 指标 2.10.0
npm
36薄弱健康指数
pavlealeksic/playwright-afp
Stop website fingerprinting techniques playwright edition
JavaScript★ 19↓ 4,152/月2026年9月2日
MIT2026年9月2日 · 指标 2.10.0
Go
36薄弱健康指数
rusq/chromedl
Go library for scraping or downloading files bypassing Cloudflare protection and browser checks
Go★ 352026年9月5日
MIT2026年9月5日 · 指标 2.10.0
PyPI
34存在风险健康指数
scrapy/parsel
Parsel lets you extract data from XML/HTML documents using XPath or CSS selectors
Python★ 1,350↓ 4.7M/月2026年8月13日
BSD-3-Clause2026年8月13日 · 指标 2.10.0
PyPI
34存在风险健康指数
scrapy/scrapy
Scrapy, a fast high-level web crawling & scraping framework for Python.
Python★ 63.8K2026年8月12日
BSD-3-Clause2026年8月12日 · 指标 2.10.0
npm
31存在风险健康指数
ddoojoang/ttj-skills-playwright
该仓库未发布描述。
TypeScript · JavaScript★ 1↓ 3,696/月2026年7月29日
无许可证2026年7月29日 · 指标 2.10.0
PyPI
31存在风险健康指数
ultrafunkamsterdam/undetected-chromedriver
Custom Selenium Chromedriver | Zero-Config | Passes ALL bot mitigation systems (like Distil / Imperva/ Datadadome / CloudFlare IUAM)
Python★ 12.8K2026年8月12日
GPL-3.02026年8月12日 · 指标 2.10.0
Maven
29存在风险健康指数
code4craft/webmagic
A scalable web crawler framework for Java.
Java · HTML★ 11.7K2026年8月27日
Apache-2.02026年8月27日 · 指标 2.10.0
Go
29存在风险健康指数
syswraith/heimdall
Fast, concurrent username OSINT tool in Go — checks presence across sites via HTTP, JSON, headless browser, and OpenGraph parsing, with SQLite caching and YAML-driven site configs.
Go★ 02026年7月24日
AGPL-3.02026年7月24日 · 指标 2.10.0
PyPI
28存在风险健康指数
VeNoMouS/cloudscraper
A Python module to bypass Cloudflare's anti-bot page.
Python★ 6,730↓ 4.5M/月2026年8月28日
MIT2026年8月28日 · 指标 2.10.0
PyPI
26存在风险健康指数
scrapedatshi/scrapedatshi-py
该仓库未发布描述。
Python★ 0↓ 4,092/月2026年7月17日
MIT2026年7月17日 · 指标 2.10.0
PyPI
21存在风险健康指数
scrapedatshi/scrapedatshi-mcp
该仓库未发布描述。
Python★ 0↓ 2,606/月2026年7月18日
MIT2026年7月18日 · 指标 2.10.0
npm
20存在风险健康指数
karthikuj/sasori
Sasori is a dynamic web crawler powered by Puppeteer, designed for lightning-fast endpoint discovery.
JavaScript★ 145↓ 78/月2026年8月4日
MIT2026年8月4日 · 指标 2.10.0
PyPI
19危急健康指数
fake-useragent/fake-useragent
Up-to-date simple useragent faker with real world database
Python★ 4,049↓ 10.8M/月2026年8月28日
Apache-2.02026年8月28日 · 指标 2.10.0
PyPI
15危急健康指数
binarynightowl/covid19_python
A fast, powerful, and flexible way to get up to date COVID-19 data for any major city, state, country, and total world wide data, with just one line of code
Python★ 11↓ 180/月2026年7月15日
MIT2026年7月15日 · 指标 2.10.0