Todas las etiquetas
Etiqueta del catálogo

#benchmark

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

73 registros
Con la etiqueta «benchmark»Ordenado por índice de salud
npm
71Buenoíndice de salud
itayinbarr/little-coder
A harness optimized to smaller LLMs
TypeScript · Python · JavaScript★ 1830↓ 3860/mes23 jul 2026
Apache-2.023 jul 2026 · métricas 2.10.0
npm
69Buenoíndice de salud
KryptSec/oasis
Open-source AI security benchmarking CLI. Measure how AI models perform offensive security tasks with MITRE ATT&CK analysis and KSM scoring.
TypeScript★ 29↓ 39/mes28 ago 2026
MIT28 ago 2026 · métricas 2.10.0
crates.io
69Buenoíndice de salud
dfinity/canbench
A benchmarking framework for canisters on the Internet Computer.
Rust★ 21↓ 16.8K/mes9 ago 2026
Apache-2.09 ago 2026 · métricas 2.10.0
PyPI
67Buenoíndice de salud
OpenAdaptAI/openadapt-evals
Evaluation infrastructure for GUI agent benchmarks
Python★ 2↓ 2941/mes28 jul 2026
MIT28 jul 2026 · métricas 2.10.0
Go
67Buenoíndice de salud
RediSearch/ftsb
Full Text Search Benchmark, a tool for comparing and evaluating full-text search engines.
Python · Go★ 2619 jul 2026
MIT19 jul 2026 · métricas 2.10.0
crates.io
67Buenoíndice de salud
criterion-rs/criterion.rs
El repositorio no publica descripción.
Rust · HTML★ 410↓ 36.6M/mes8 ago 2026
Apache-2.08 ago 2026 · métricas 2.10.0
67Buenoíndice de salud
dadhi/FastExpressionCompiler
Fast Compiler for C# Expression Trees and the lightweight LightExpression alternative. Diagnostic and code generation tools for the expressions.
C#★ 137018 jul 2026
MIT18 jul 2026 · métricas 2.10.0
crates.io
67Buenoíndice de salud
imazen/zenbench
Interleaved microbenchmarking for Rust — paired statistics, CI regression testing, criterion migration
Rust★ 3↓ 6495/mes17 jul 2026
Licencia propia17 jul 2026 · métricas 2.10.0
PyPI · crates.io
63Moderadoíndice de salud
FedericoTs/quantprobe
Run a 110B on a 2016 PC with 16 GB RAM. Know your tok/s before you download. Placement beats budget: predicts speed + memory fit for any GGUF on your exact hardware, self-calibrates, emits the exact llama.cpp command — or 'quantprobe auto' does it all. Falsification-tested laws; misses published at full size. pip install quantprobe
Python★ 86↓ 6494/mes20 ago 2026
MIT20 ago 2026 · métricas 2.10.0
PyPI
63Moderadoíndice de salud
druide67/asiai
Multi-engine LLM benchmark & monitoring CLI for Apple Silicon
Python★ 11↓ 4662/mes18 jul 2026
Apache-2.018 jul 2026 · métricas 2.10.0
Go
63Moderadoíndice de salud
filipecosta90/ftsb
Full Text Search Benchmark, a tool for comparing and evaluating full-text search engines.
Python · Go★ 2617 jul 2026
MIT17 jul 2026 · métricas 2.10.0
PyPI
63Moderadoíndice de salud
google-ai-edge/LiteRT-CLI
A convenient CLI to streamline LiteRT related development workflows, including converting, quantizing, compiling, managing, running, benchmarking and visualizing LiteRT (TFLite) models on various hardwares (CPU / GPU / NPU) across platforms (desktop, mobile or cloud).
Python★ 34↓ 80/mes15 jul 2026
Apache-2.015 jul 2026 · métricas 2.10.0
npm
63Moderadoíndice de salud
mgechev/skillgrade
"Unit tests" for your agent skills
TypeScript★ 661↓ 2290/mes5 ago 2026
MIT5 ago 2026 · métricas 2.10.0
Hex
62Moderadoíndice de salud
aryaminus/controlkeel
Agent control plane for governed AI coding: validate changes, enforce policy gates, track findings, proofs, and evals based on your habits.
Elixir★ 1017 jul 2026
Licencia propia17 jul 2026 · métricas 2.10.0
crates.io
62Moderadoíndice de salud
kornelski/dssim
Image similarity comparison simulating human perception (multiscale SSIM in Rust)
Rust★ 1194↓ 28.6K/mes1 sept 2026
AGPL-3.01 sept 2026 · métricas 2.10.0
PyPI · crates.io
59Moderadoíndice de salud
apitap/apitap-lib
Move whole tables between databases fast — Postgres, MySQL, ClickHouse, BigQuery. Rust engine, one-line Python API, bounded memory.
Rust · Python★ 5018 ago 2026
MIT18 ago 2026 · métricas 2.10.0
PyPI
59Moderadoíndice de salud
vlbthambawita/ECGBench
Reproduciable ECG Benchmark data from Open access datasets
Python★ 3↓ 2288/mes6 ago 2026
MIT6 ago 2026 · métricas 2.10.0
npm
57Moderadoíndice de salud
anhldh/r3f-monitor
An easy tool to monitor the performance of your R3F application.
TypeScript · CSS★ 8↓ 2663/mes18 jul 2026
MIT18 jul 2026 · métricas 2.10.0
crates.io
56Moderadoíndice de salud
THeK3nger/movingai-rust
Map/Scenario parser for the MovingAI benchmark format
Roff★ 9↓ 2286/mes22 ago 2026
MIT22 ago 2026 · métricas 2.10.0
PyPI
56Moderadoíndice de salud
a-r-j/ProteinWorkshop
Benchmarking framework for protein representation learning. Includes a large number of pre-training and downstream task datasets, models and training/task utilities. (ICLR 2024)
Python · Jupyter Notebook★ 27515 jul 2026
MIT15 jul 2026 · métricas 2.10.0
npm
56Moderadoíndice de salud
onury/perfy
A tiny, zero-dependency utility for measuring code execution time in high-resolution real time. Works in Node.js, browsers, Deno and Bun.
TypeScript★ 55↓ 143.7K/mes1 ago 2026
MIT1 ago 2026 · métricas 2.10.0
PyPI
54Moderadoíndice de salud
RRZE-HPC/kerncraft
Loop Kernel Analysis and Performance Modeling Toolkit
Jupyter Notebook · Python★ 9911 ago 2026
AGPL-3.011 ago 2026 · métricas 2.10.0
PyPI
54Moderadoíndice de salud
danaug23/harness-arena
Your model, many harnesses, many benchmarks.
Python · HTML★ 2↓ 2988/mes23 ago 2026
Apache-2.023 ago 2026 · métricas 2.10.0
npm
54Moderadoíndice de salud
inferock/inferock-bench
Local LLM cost-tracking proxy for OpenAI, Anthropic, Gemini, and pinned OpenRouter calls with token usage, failure, and billing-integrity receipts.
TypeScript★ 123↓ 6119/mes29 jul 2026
Licencia propia29 jul 2026 · métricas 2.10.0
Packagist
54Moderadoíndice de salud
infocyph/PHPForge
Shared Composer-powered QA, refactoring, benchmark, release, hook and CI tooling for PHP projects.
PHP★ 1↓ 2564/mes3 ago 2026
MIT3 ago 2026 · métricas 2.10.0
RubyGems
54Moderadoíndice de salud
michaelherold/benchmark-memory
Memory profiling benchmark style, for Ruby 2.1+
Ruby★ 25020 jul 2026
MIT20 jul 2026 · métricas 2.10.0
RubyGems
53Moderadoíndice de salud
pboling/gem_bench
🪑 Benchmark different versions of same or similar gems & Static Gemfile and installed gem library source code analysis
Ruby★ 9417 jul 2026
MIT17 jul 2026 · métricas 2.10.0
crates.io
51Moderadoíndice de salud
beling/bsuccinct-rs
Rust libraries and programs focused on succinct data structures
Rust★ 171↓ 343.7K/mes19 ago 2026
Apache-2.019 ago 2026 · métricas 2.10.0
PyPI
51Moderadoíndice de salud
sablier-ai/finval
Rigorous validation for synthetic financial time series — 19 financial stylized-fact metrics, used as the FinBench scoring backend
Python★ 1↓ 344/mes31 jul 2026
MIT31 jul 2026 · métricas 2.10.0
PyPI
50Moderadoíndice de salud
SidRichardsQuantum/Quantum_Backend_Bench
Run the same quantum circuit across multiple backends and compare performance, depth, gate counts, and noise effects in a single, unified workflow.
Python · Jupyter Notebook★ 1↓ 534/mes20 ago 2026
MIT20 ago 2026 · métricas 2.10.0