Todas las etiquetas
Etiqueta del catálogo

#token-optimization

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

42 registros
Con la etiqueta «token-optimization»Ordenado por índice de salud
npm · crates.io · PyPI
94Excepcionalíndice de salud
headroomlabs-ai/headroom
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
Python · Rust★ 64.8K↓ 115.9K/mes4 ago 2026
Apache-2.04 ago 2026 · métricas 2.10.0
crates.io · npm
93Excepcionalíndice de salud
yvgude/lean-ctx
Control what your AI can see. LeanCTX (Lean Context) is the context intelligence layer for AI agents — one local Rust binary that decides what they read, remembers what they learn, guards what they touch, and proves what they save. 60–90% fewer tokens as the receipt. 76 MCP tools, 30+ agents, local-first.
Rust★ 3631↓ 3190/mes22 ago 2026
Apache-2.022 ago 2026 · métricas 2.10.0
crates.io · npm
90Excelenteíndice de salud
rtk-ai/rtk
CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies
Rust★ 74.7K4 ago 2026
Apache-2.04 ago 2026 · métricas 2.10.0
PyPI
86Excelenteíndice de salud
jgravelle/jdocmunch-mcp
The leading, most token-efficient MCP server for documentation exploration and retrieval via structured section indexing
Python★ 203↓ 19.1K/mes22 ago 2026
Licencia propia22 ago 2026 · métricas 2.10.0
npm
86Excelenteíndice de salud
ooples/token-optimizer-mcp
Intelligent token optimization for Claude Code - achieving 95%+ token reduction through caching, compression, and smart tool intelligence
TypeScript★ 454↓ 2292/mes29 jul 2026
MIT29 jul 2026 · métricas 2.10.0
npm
84Excelenteíndice de salud
teamchong/pxpipe
cut Claude Code token usage by rendering text context as images
TypeScript★ 7291↓ 9307/mes28 ago 2026
MIT28 ago 2026 · métricas 2.10.0
PyPI · npm
83Excelenteíndice de salud
jgravelle/jcodemunch-mcp
Cut AI token costs 95%+ on code exploration. The leading MCP server for precise, symbol-level GitHub code retrieval via tree-sitter AST. Works with Claude Code, Cursor & any MCP client. 313B+ tokens saved.
Python★ 2024↓ 88.4K/mes16 jul 2026
Licencia propia16 jul 2026 · métricas 2.10.0
Go
81Excelenteíndice de salud
edouard-claude/snip
CLI proxy that reduces LLM token usage by 60-90%. Declarative YAML filters for Claude Code, Cursor, Copilot, Gemini. rtk alternative in Go.
Go★ 37419 jul 2026
MIT19 jul 2026 · métricas 2.10.0
PyPI · npm · crates.io
80Excelenteíndice de salud
juyterman1000/entroly
Auditable context engineering for AI agents: context optimization, recoverable context compression, receipts, answer verification, and MCP for Claude Code, Codex, OpenClaw.
Python · Rust★ 428↓ 14.6K/mes21 jul 2026
Apache-2.021 jul 2026 · métricas 2.10.0
npm · crates.io
78Buenoíndice de salud
dPeluChe/trs
Token-Reducing Shell — terminal output compression for AI coding agents
Rust★ 11↓ 857/mes5 sept 2026
MIT5 sept 2026 · métricas 2.10.0
PyPI · npm
77Buenoíndice de salud
SonAIengine/graph-tool-call
Graph-based tool retrieval for LLM agents — 248 tools → 82% accuracy, 79% fewer tokens. Zero dependencies. OpenAPI / MCP / LangChain.
Python★ 7↓ 2636/mes1 ago 2026
MIT1 ago 2026 · métricas 2.10.0
crates.io · npm
75Buenoíndice de salud
ratel-ai/ratel
Context engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.
Rust · Python · TypeScript★ 23320 jul 2026
MIT20 jul 2026 · métricas 2.10.0
npm
73Buenoíndice de salud
cytostack/openwolf
Sharper context. Fewer tokens. Open-source middleware for Claude Code.
TypeScript★ 2129↓ 10.7K/mes17 jul 2026
AGPL-3.017 jul 2026 · métricas 2.10.0
PyPI · crates.io
73Buenoíndice de salud
fkiene/llmtrim
Local proxy that compresses your LLM API requests so you pay less, with no change to the answers. Trims wasted tokens from prompts, history, tool output, and code before they're sent: -31% input / -74% output, measured live. Any provider, no extra model calls. Also an MCP server and embeddable library (Rust, Python, Ruby, Kotlin, Swift, JS/TS).
Rust★ 179↓ 5878/mes26 jul 2026
MPL-2.026 jul 2026 · métricas 2.10.0
PyPI
69Buenoíndice de salud
maheshmakvana/graphsift
Token Saver for Claude, GPT-5 & Gemini. 80-150x code context reduction, F1 0.85. AST dependency graph, ranked context selection, 19 CLI compressors, MCP server, agent memory. Save LLM tokens — zero telemetry.
Python★ 4↓ 2797/mes4 ago 2026
MIT4 ago 2026 · métricas 2.10.0
Go
69Buenoíndice de salud
sphragis-oss/isthmos
Local context-compression layer for agent tool outputs. Claude Code PostToolUse hook or generic filter, single Go binary.
Go★ 031 jul 2026
Apache-2.031 jul 2026 · métricas 2.10.0
npm
67Buenoíndice de salud
wundercorp/openmodel
Use any AI model, locally or in the cloud, while tracking on-device token usage and collecting token telemetry data.
TypeScript · JavaScript · CSS★ 18↓ 3140/mes7 ago 2026
Apache-2.07 ago 2026 · métricas 2.10.0
npm · crates.io · PyPI
62Moderadoíndice de salud
Tura-AI/tura
Across 348 long-horizon benchmark sessions, Tura used up to 83.1% fewer turns on the rewrite benchmark and improved the DeepSWE pass rate by up to 16.7 percentage points compared with Codex CLI.
Rust · TypeScript · JavaScript★ 56↓ 3378/mes16 jul 2026
AGPL-3.016 jul 2026 · métricas 2.10.0
PyPI
60Moderadoíndice de salud
Chuzom/Chuzom
Lightweight signal-driven LLM router for Claude Code, Cursor, Codex, Gemini CLI, and Codex CLI
Python★ 19↓ 2660/mes1 ago 2026
MIT1 ago 2026 · métricas 2.10.0
npm
59Moderadoíndice de salud
TaewoooPark/Agent-Blackbox
Local-first flight recorder for coding agents : replay every run as a live session map, score the context bill, and write the fix back into AGENTS.md — no API key, one npx command.
TypeScript★ 57↓ 3988/mes21 jul 2026
MIT21 jul 2026 · métricas 2.10.0
Packagist · npm
59Moderadoíndice de salud
mischasigtermans/laravel-toon
TOON encoding for Laravel. Encode data for AI/LLMs with ~50% fewer tokens than JSON.
PHP★ 144↓ 11K/mes23 jul 2026
MIT23 jul 2026 · métricas 2.10.0
npm
59Moderadoíndice de salud
psjostrom/frontload
Local-first context gateway that helps AI coding agents read less, spend less, and stay grounded in your repo.
TypeScript★ 0↓ 2191/mes24 jul 2026
Sin licencia24 jul 2026 · métricas 2.10.0
npm
59Moderadoíndice de salud
syntaxPriest/openvisio-oss
El repositorio no publica descripción.
TypeScript★ 10↓ 2951/mes23 jul 2026
MIT23 jul 2026 · métricas 2.10.0
npm
57Moderadoíndice de salud
bacnh85/pi-extensions
Pi coding agent extension packages — tools, skills, and integrations.
TypeScript★ 9↓ 7997/mes29 ago 2026
Sin licencia29 ago 2026 · métricas 2.10.0
npm · Go · PyPI
57Moderadoíndice de salud
firstops-dev/whittle
Carves your agent's tool outputs down to what matters. Never cuts what doesn't come back.
Go · Python★ 61↓ 37/mes5 sept 2026
Apache-2.05 sept 2026 · métricas 2.10.0
npm
56Moderadoíndice de salud
darkrei08/Wizard-AI
El repositorio no publica descripción.
HTML · Shell★ 1↓ 6010/mes18 jul 2026
AGPL-3.018 jul 2026 · métricas 2.10.0
npm
56Moderadoíndice de salud
kitepon-rgb/aiterm-mcp
One persistent MCP terminal your AI drives — and launches other coding agents (Codex/Grok/Composer) into. SSH, containers, and REPLs nest as text you send in. tmux-backed, token-reduced reads, headless over MCP.
JavaScript · TypeScript · Python★ 1↓ 2175/mes15 jul 2026
MIT15 jul 2026 · métricas 2.10.0
Go
54Moderadoíndice de salud
blackwell-systems/gcf-go
GCF Go implementation. 100% LLM comprehension on every frontier model. 50-92% fewer tokens than JSON. 43B+ round-trips verified. Zero dependencies.
Go★ 422 ago 2026
MIT22 ago 2026 · métricas 2.10.0
PyPI
54Moderadoíndice de salud
wesleysimplicio/simplicio-loop
🔁 Finishes your entire backlog while you sleep. The AI orchestrator that DOES the work end-to-end on ANY LLM — discover → implement → verify → merge → 24/7 — behind safety gates, at up to 90% fewer tokens. 48 extension points. Not a chatbot. A worker.
Python★ 10↓ 3430/mes31 jul 2026
MIT31 jul 2026 · métricas 2.10.0
npm
51Moderadoíndice de salud
bassprofressor-lab/openwolf-enhanced
Enhanced fork of OpenWolf — a token-conscious second brain for Claude Code, with bounded storage, self-maintenance (openwolf doctor), .wolfignore scoping, and tunable retention. AGPL-3.0.
TypeScript · JavaScript★ 6↓ 5230/mes30 jul 2026
AGPL-3.030 jul 2026 · métricas 2.10.0