Go
wentbackward/llm-proxy
A superfast proxy and smart load-balancer for AI Inference — virtualize models, share local and provider back-ends, optimal caching, lock-in sampling parameters, debug message flow and obtain OTel metrics
MITJul 17, 2026 · metrics 2.10.0