Todas las etiquetas
Etiqueta del catálogo

#benchmark

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

73 registros
Con la etiqueta «benchmark»Ordenado por índice de salud
PyPI
95Excepcionalíndice de salud
embeddings-benchmark/mteb
MTEB: State-of-the-art evaluation of embeddings across languages and modalities
Python · Jupyter Notebook★ 336219 jul 2026
Apache-2.019 jul 2026 · métricas 2.10.0
npm
95Excepcionalíndice de salud
tinylibs/tinybench
🔎 A simple, tiny and lightweight benchmarking library!
TypeScript★ 2354↓ 296M/mes4 ago 2026
MIT4 ago 2026 · métricas 2.10.0
NuGet
94Excepcionalíndice de salud
dotnet/BenchmarkDotNet
Powerful .NET library for benchmarking
C#★ 11.5K27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
npm
93Excepcionalíndice de salud
artilleryio/artillery
The complete load testing platform. Everything you need for production-grade load tests. Serverless & distributed. Load test with Playwright. Load test HTTP APIs, GraphQL, WebSocket, and more. Use any Node.js module.
TypeScript · JavaScript★ 9059↓ 1.1M/mes28 ago 2026
MPL-2.028 ago 2026 · métricas 2.10.0
npm · PyPI
91Excelenteíndice de salud
joshuaswarren/remnic
Open-source memory and context for user-aware agents: scoped memory, provenance, retrieval quality, correction, boundaries, evals, and MCP/HTTP access.
TypeScript★ 176↓ 203.7K/mes22 ago 2026
MIT22 ago 2026 · métricas 2.10.0
PyPI
90Excelenteíndice de salud
airspeed-velocity/asv
Airspeed Velocity: A simple Python benchmarking tool with web-based reporting
Python · JavaScript★ 1008↓ 508.4K/mes21 jul 2026
BSD-3-Clause21 jul 2026 · métricas 2.10.0
PyPI · npm
89Excelenteíndice de salud
UiPath/coder_eval
Test that your Claude Code skills, MCP servers, and CLIs actually work when an agent uses them — sandboxed YAML suites, activation checks, A/B experiments, CI gates.
Python · TypeScript★ 116↓ 13.9K/mes19 ago 2026
Apache-2.019 ago 2026 · métricas 2.10.0
PyPI
89Excelenteíndice de salud
fitbenchmarking/fitbenchmarking
Tool for comparing the run time and accuracy of minimizers on fit benchmarking problems
Python★ 1625 jul 2026
BSD-3-Clause25 jul 2026 · métricas 2.10.0
Go
89Excelenteíndice de salud
gofiber/utils
:zap: A collection of common functions for Fiber with better performance, fewer allocations, and fewer dependencies.
Go★ 5722 ago 2026
MIT22 ago 2026 · métricas 2.10.0
crates.io
89Excelenteíndice de salud
gungraun/gungraun
High-precision, one-shot and consistent benchmarking framework/harness for Rust. All Valgrind tools at your fingertips.
Rust · Roff★ 317↓ 412.5K/mes22 ago 2026
Apache-2.022 ago 2026 · métricas 2.10.0
PyPI
87Excelenteíndice de salud
SWE-bench/SWE-bench
SWE-bench: Can Language Models Resolve Real-world Github Issues?
Python★ 5726↓ 46.4M/mes28 ago 2026
MIT28 ago 2026 · métricas 2.10.0
crates.io
87Excelenteíndice de salud
pawurb/hotpath-rs
Quickly find bottlenecks in Rust - one profiler for CPU, memory, SQL, HTTP, I/O and async code.
Rust★ 1688↓ 174.3K/mes22 ago 2026
MIT22 ago 2026 · métricas 2.10.0
PyPI
86Excelenteíndice de salud
MichaelGrupp/evo
Python package for the evaluation of odometry and SLAM
Python★ 4309↓ 210.7K/mes28 ago 2026
GPL-3.028 ago 2026 · métricas 2.10.0
crates.io · Maven · npm
86Excelenteíndice de salud
SaaSy-Solutions/mockforge
Comprehensive mocking framework for REST, gRPC, GraphQL & WebSockets. Features intelligent RAG-driven data synthesis, latency/fault injection, HTTP bridge, WASM plugins, E2E encryption, workspace sync, and modern admin UI. Production-ready with Docker support.
Rust · TypeScript · HTML★ 11↓ 9833/mes5 sept 2026
Licencia propia5 sept 2026 · métricas 2.10.0
PyPI · npm
86Excelenteíndice de salud
robocurve/inspect-robots
Open source evals for physical AI. Run any LLM/VLA on any arm/humanoid against any real/sim benchmark.
Python★ 269↓ 5131/mes6 sept 2026
MIT6 sept 2026 · métricas 2.10.0
PyPI
84Excelenteíndice de salud
CodSpeedHQ/pytest-codspeed
A pytest plugin to create benchmarks
Python★ 134↓ 2.4M/mes27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
npm
84Excelenteíndice de salud
vava-nessa/free-coding-models
Find, benchmark and install in CLI 170+ FREE coding LLM models across 15+ providers in real time
HTML · JavaScript★ 2186↓ 5344/mes23 jul 2026
Licencia propia23 jul 2026 · métricas 2.10.0
PyPI
83Excelenteíndice de salud
SWE-bench/SWE-smith
[NeurIPS 2025 D&B Spotlight] Scaling Data for SWE-agents
Python★ 752↓ 16.8M/mes27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
crates.io
81Excelenteíndice de salud
hatoo/oha
Ohayou(おはよう), HTTP load generator, inspired by rakyll/hey with tui animation.
Rust★ 10.5K↓ 1822/mes28 ago 2026
MIT28 ago 2026 · métricas 2.10.0
crates.io
81Excelenteíndice de salud
lance0/xfr
A modern iperf3 alternative with a live TUI, multi-client server, and QUIC support. Built in Rust.
Rust★ 517↓ 179/mes23 jul 2026
Apache-2.023 jul 2026 · métricas 2.10.0
crates.io
80Excelenteíndice de salud
bencherdev/bencher
🐰 Bencher - Continuous Benchmarking
Rust · MDX★ 874↓ 11/mes27 jul 2026
Licencia propia27 jul 2026 · métricas 2.10.0
PyPI
80Excelenteíndice de salud
benchflow-ai/benchflow
Research infra for creating RL environments, post-training, and evals.
Python★ 317↓ 6133/mes9 ago 2026
Apache-2.09 ago 2026 · métricas 2.10.0
npm · Go
80Excelenteíndice de salud
goptics/vizb
A tabular data visualization engine from your local to CI/CD pipeline
Go · TypeScript★ 79↓ 4090/mes17 jul 2026
MIT17 jul 2026 · métricas 2.10.0
Go
80Excelenteíndice de salud
oneclickvirt/ecs
VPS Fusion Monster Server Test GO Version Aiming to be the most comprehensive server testing project, implemented in Go with zero environment dependencies. VPS融合怪服务器测评项目 GO版本 尽量成为最全能的服务器测评项目,使用 Go 实现,无需任何环境依赖。
Go★ 226429 jul 2026
GPL-3.029 jul 2026 · métricas 2.10.0
PyPI
78Buenoíndice de salud
Sahil170595/Chimeraforge
PyPI capacity-planning CLI for LLM deployment. pip install chimeraforge.
Python★ 2↓ 2506/mes22 ago 2026
MIT22 ago 2026 · métricas 2.10.0
75Buenoíndice de salud
soenneker/soenneker.tests.benchmark
An abstract class for benchmarking tests in .NET, integrating BenchmarkDotNet with Xunit's output helper and providing a method to log benchmark summaries asynchronously.
C#★ 016 jul 2026
MIT16 jul 2026 · métricas 2.10.0
73Buenoíndice de salud
inikep/lzbench
lzbench is an in-memory benchmark of open-source compressors
C · C++★ 107421 jul 2026
Licencia propia21 jul 2026 · métricas 2.10.0
crates.io · PyPI
73Buenoíndice de salud
sebastienrousseau/http-handle
Lightweight Rust HTTP server library. Sync + async, HTTP/1.1 keep-alive, HTTP/2 (h2c), HTTP/3 ALPN. 180 k req/s on Linux/arm64.
Rust★ 1↓ 2031/mes28 ago 2026
Apache-2.028 ago 2026 · métricas 2.10.0
PyPI
71Buenoíndice de salud
ionelmc/pytest-benchmark
pytest fixture for benchmarking code
Python★ 144417 jul 2026
BSD-2-Clause17 jul 2026 · métricas 2.10.0