全部标签
目录标签

#benchmark

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

73 条记录
标签为“benchmark”按健康指数排序
PyPI
95卓越健康指数
embeddings-benchmark/mteb
MTEB: State-of-the-art evaluation of embeddings across languages and modalities
Python · Jupyter Notebook★ 3,3622026年7月19日
Apache-2.02026年7月19日 · 指标 2.10.0
npm
95卓越健康指数
tinylibs/tinybench
🔎 A simple, tiny and lightweight benchmarking library!
TypeScript★ 2,354↓ 296M/月2026年8月4日
MIT2026年8月4日 · 指标 2.10.0
NuGet
94卓越健康指数
dotnet/BenchmarkDotNet
Powerful .NET library for benchmarking
C#★ 11.5K2026年8月27日
MIT2026年8月27日 · 指标 2.10.0
npm
93卓越健康指数
artilleryio/artillery
The complete load testing platform. Everything you need for production-grade load tests. Serverless & distributed. Load test with Playwright. Load test HTTP APIs, GraphQL, WebSocket, and more. Use any Node.js module.
TypeScript · JavaScript★ 9,059↓ 1.1M/月2026年8月28日
MPL-2.02026年8月28日 · 指标 2.10.0
npm · PyPI
91优秀健康指数
joshuaswarren/remnic
Open-source memory and context for user-aware agents: scoped memory, provenance, retrieval quality, correction, boundaries, evals, and MCP/HTTP access.
TypeScript★ 176↓ 203.7K/月2026年8月22日
MIT2026年8月22日 · 指标 2.10.0
PyPI
90优秀健康指数
airspeed-velocity/asv
Airspeed Velocity: A simple Python benchmarking tool with web-based reporting
Python · JavaScript★ 1,008↓ 508.4K/月2026年7月21日
BSD-3-Clause2026年7月21日 · 指标 2.10.0
PyPI · npm
89优秀健康指数
UiPath/coder_eval
Test that your Claude Code skills, MCP servers, and CLIs actually work when an agent uses them — sandboxed YAML suites, activation checks, A/B experiments, CI gates.
Python · TypeScript★ 116↓ 13.9K/月2026年8月19日
Apache-2.02026年8月19日 · 指标 2.10.0
PyPI
89优秀健康指数
fitbenchmarking/fitbenchmarking
Tool for comparing the run time and accuracy of minimizers on fit benchmarking problems
Python★ 162026年7月25日
BSD-3-Clause2026年7月25日 · 指标 2.10.0
Go
89优秀健康指数
gofiber/utils
:zap: A collection of common functions for Fiber with better performance, fewer allocations, and fewer dependencies.
Go★ 572026年8月22日
MIT2026年8月22日 · 指标 2.10.0
crates.io
89优秀健康指数
gungraun/gungraun
High-precision, one-shot and consistent benchmarking framework/harness for Rust. All Valgrind tools at your fingertips.
Rust · Roff★ 317↓ 412.5K/月2026年8月22日
Apache-2.02026年8月22日 · 指标 2.10.0
PyPI
87优秀健康指数
SWE-bench/SWE-bench
SWE-bench: Can Language Models Resolve Real-world Github Issues?
Python★ 5,726↓ 46.4M/月2026年8月28日
MIT2026年8月28日 · 指标 2.10.0
crates.io
87优秀健康指数
pawurb/hotpath-rs
Quickly find bottlenecks in Rust - one profiler for CPU, memory, SQL, HTTP, I/O and async code.
Rust★ 1,688↓ 174.3K/月2026年8月22日
MIT2026年8月22日 · 指标 2.10.0
PyPI
86优秀健康指数
MichaelGrupp/evo
Python package for the evaluation of odometry and SLAM
Python★ 4,309↓ 210.7K/月2026年8月28日
GPL-3.02026年8月28日 · 指标 2.10.0
crates.io · Maven · npm
86优秀健康指数
SaaSy-Solutions/mockforge
Comprehensive mocking framework for REST, gRPC, GraphQL & WebSockets. Features intelligent RAG-driven data synthesis, latency/fault injection, HTTP bridge, WASM plugins, E2E encryption, workspace sync, and modern admin UI. Production-ready with Docker support.
Rust · TypeScript · HTML★ 11↓ 9,833/月2026年9月5日
自定义许可证2026年9月5日 · 指标 2.10.0
PyPI · npm
86优秀健康指数
robocurve/inspect-robots
Open source evals for physical AI. Run any LLM/VLA on any arm/humanoid against any real/sim benchmark.
Python★ 269↓ 5,131/月2026年9月6日
MIT2026年9月6日 · 指标 2.10.0
PyPI
84优秀健康指数
CodSpeedHQ/pytest-codspeed
A pytest plugin to create benchmarks
Python★ 134↓ 2.4M/月2026年8月27日
MIT2026年8月27日 · 指标 2.10.0
npm
84优秀健康指数
vava-nessa/free-coding-models
Find, benchmark and install in CLI 170+ FREE coding LLM models across 15+ providers in real time
HTML · JavaScript★ 2,186↓ 5,344/月2026年7月23日
自定义许可证2026年7月23日 · 指标 2.10.0
PyPI
83优秀健康指数
SWE-bench/SWE-smith
[NeurIPS 2025 D&B Spotlight] Scaling Data for SWE-agents
Python★ 752↓ 16.8M/月2026年8月27日
MIT2026年8月27日 · 指标 2.10.0
crates.io
81优秀健康指数
hatoo/oha
Ohayou(おはよう), HTTP load generator, inspired by rakyll/hey with tui animation.
Rust★ 10.5K↓ 1,822/月2026年8月28日
MIT2026年8月28日 · 指标 2.10.0
crates.io
81优秀健康指数
lance0/xfr
A modern iperf3 alternative with a live TUI, multi-client server, and QUIC support. Built in Rust.
Rust★ 517↓ 179/月2026年7月23日
Apache-2.02026年7月23日 · 指标 2.10.0
crates.io
80优秀健康指数
bencherdev/bencher
🐰 Bencher - Continuous Benchmarking
Rust · MDX★ 874↓ 11/月2026年7月27日
自定义许可证2026年7月27日 · 指标 2.10.0
PyPI
80优秀健康指数
benchflow-ai/benchflow
Research infra for creating RL environments, post-training, and evals.
Python★ 317↓ 6,133/月2026年8月9日
Apache-2.02026年8月9日 · 指标 2.10.0
npm · Go
80优秀健康指数
goptics/vizb
A tabular data visualization engine from your local to CI/CD pipeline
Go · TypeScript★ 79↓ 4,090/月2026年7月17日
MIT2026年7月17日 · 指标 2.10.0
Go
80优秀健康指数
oneclickvirt/ecs
VPS Fusion Monster Server Test GO Version Aiming to be the most comprehensive server testing project, implemented in Go with zero environment dependencies. VPS融合怪服务器测评项目 GO版本 尽量成为最全能的服务器测评项目,使用 Go 实现,无需任何环境依赖。
Go★ 2,2642026年7月29日
GPL-3.02026年7月29日 · 指标 2.10.0
PyPI
78良好健康指数
Sahil170595/Chimeraforge
PyPI capacity-planning CLI for LLM deployment. pip install chimeraforge.
Python★ 2↓ 2,506/月2026年8月22日
MIT2026年8月22日 · 指标 2.10.0
75良好健康指数
soenneker/soenneker.tests.benchmark
An abstract class for benchmarking tests in .NET, integrating BenchmarkDotNet with Xunit's output helper and providing a method to log benchmark summaries asynchronously.
C#★ 02026年7月16日
MIT2026年7月16日 · 指标 2.10.0
73良好健康指数
inikep/lzbench
lzbench is an in-memory benchmark of open-source compressors
C · C++★ 1,0742026年7月21日
自定义许可证2026年7月21日 · 指标 2.10.0
crates.io · PyPI
73良好健康指数
sebastienrousseau/http-handle
Lightweight Rust HTTP server library. Sync + async, HTTP/1.1 keep-alive, HTTP/2 (h2c), HTTP/3 ALPN. 180 k req/s on Linux/arm64.
Rust★ 1↓ 2,031/月2026年8月28日
Apache-2.02026年8月28日 · 指标 2.10.0
PyPI
71良好健康指数
ionelmc/pytest-benchmark
pytest fixture for benchmarking code
Python★ 1,4442026年7月17日
BSD-2-Clause2026年7月17日 · 指标 2.10.0