全部标签
目录标签

#kv-cache

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

5 条记录
标签为“kv-cache”按健康指数排序
PyPI
63中等健康指数
varjoranta/turboquant-vllm
TurboQuant+ KV cache compression for vLLM. 3.8x smaller KV cache, same conversation quality. Fused CUDA kernels with automatic PyTorch fallback.
Python · C++ · Cuda★ 76↓ 960/月2026年7月22日
MIT2026年7月22日 · 指标 1.13.0
PyPI
60中等健康指数
jagmarques/nexusquant
Training-free KV cache compression via E8 lattice VQ. 2-bit KV that preserves retrieval (30/30 NIAH vs TurboQuant 0/30). Calibration-free, 9 architectures validated.
Python★ 252026年7月16日
自定义许可证2026年7月16日 · 指标 1.13.0
Go · PyPI
59中等健康指数
anthony-chaudhary/fak
fak — the Fused Agent Kernel: one Go binary for AI agent loops. Wrap Claude Code/Codex/Cursor, keep long sessions cache-efficient, route work per call, run local GGUF models, and adjudicate tool calls.
Go · Python★ 152026年7月17日
Apache-2.02026年7月17日 · 指标 1.13.0
npm · PyPI
56中等健康指数
rajveer43/VeloxQuant-MLX
TurboQuant MLX implementation for Apple Silicon Faster KV cache quantization optimized for MLX
Python★ 10↓ 4,108/月2026年7月14日
MIT2026年7月14日 · 指标 1.13.0
crates.io
39存在风险健康指数
RecursiveIntell/turbo-quant
Rust implementation of TurboQuant, PolarQuant, and QJL — zero-overhead vector quantization for semantic search and KV cache compression (ICLR 2026)
Rust★ 29↓ 2,256/月2026年7月13日
MIT2026年7月13日 · 指标 1.13.0