Todas las etiquetas
Etiqueta del catálogo

#llm-observability

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

20 registros
Con la etiqueta «llm-observability»Ordenado por índice de salud
PyPI · npm
99Excepcionalíndice de salud
pydantic/logfire
AI observability platform for production LLM and agent systems.
Python★ 4442↓ 14.6M/mes27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
Go · npm
97Excepcionalíndice de salud
maximhq/bifrost
Fastest enterprise AI gateway (50x faster than LiteLLM) with adaptive load balancer, cluster mode, guardrails, 1000+ models support & <100 µs overhead at 5k RPS.
Go · TypeScript★ 760828 ago 2026
Apache-2.028 ago 2026 · métricas 2.10.0
npm · PyPI
96Excepcionalíndice de salud
comet-ml/opik
Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
Python · TypeScript★ 21.1K↓ 111.9K/mes5 ago 2026
Apache-2.05 ago 2026 · métricas 2.10.0
npm
94Excepcionalíndice de salud
langfuse/langfuse
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23
TypeScript★ 32.5K5 ago 2026
Licencia propia5 ago 2026 · métricas 2.10.0
npm
93Excepcionalíndice de salud
latitude-dev/latitude-llm
Latitude is the open-source AI monitoring platform.
TypeScript · Python★ 460428 ago 2026
MIT28 ago 2026 · métricas 2.10.0
PyPI
92Excelenteíndice de salud
JudgmentLabs/judgeval
The Continuous-Improvement Stack for Agents. Our environment data and evals power agent improvement and monitoring.
Python★ 1057↓ 233.1K/mes13 ago 2026
Apache-2.013 ago 2026 · métricas 2.10.0
PyPI · npm
81Excelenteíndice de salud
douglasmonsky/codex-usage-tracker
Local-first MCP tools and dashboard for investigating Codex token usage, credits, costs, caching, and thread patterns.
Python · TypeScript★ 187↓ 5494/mes25 jul 2026
MIT25 jul 2026 · métricas 2.10.0
PyPI
80Excelenteíndice de salud
Mandark-droid/genai_otel_instrument
GenAI OpenTelemetry Auto-Instrumentation Library A comprehensive wrapper for automatic instrumentation of LLM/GenAI applications Supports all major LLM providers and MCP (Model Context Protocol) tool calls
Python★ 3↓ 4343/mes5 sept 2026
Apache-2.05 sept 2026 · métricas 2.10.0
PyPI · npm
80Excelenteíndice de salud
ShekharBhardwaj/AgenticLedger
Local-first flight recorder for AI agents: every LLM call, tool call, cost, and loop captured by a transparent proxy. Zero code changes.
Python · TypeScript★ 5↓ 4061/mes23 ago 2026
MIT23 ago 2026 · métricas 2.10.0
Go
80Excelenteíndice de salud
nudgebee/node-agent
Per-node observability agent for Kubernetes and Linux hosts. Gathers container and host metrics, logs, and L7 traffic via eBPF; exports to Prometheus and OpenTelemetry. Includes LLM API observability.
Go★ 022 jul 2026
Apache-2.022 jul 2026 · métricas 2.10.0
PyPI
77Buenoíndice de salud
alizahidraja/isnad
Grade every agent, scraper and model in a claim's chain — provenance, trust scoring and audit evidence for LLM pipelines
Python★ 37↓ 4204/mes29 ago 2026
Apache-2.029 ago 2026 · métricas 2.10.0
PyPI
73Buenoíndice de salud
phierceweb/pf-core
Python foundation for LLM apps whose prompts and spend you can actually see — versioned prompts, every call recorded and replayable, budgets, evals, jobs.
Python★ 2↓ 1160/mes5 sept 2026
MIT5 sept 2026 · métricas 2.10.0
npm
71Buenoíndice de salud
cendorhq/cendor-sdk-js
A governed agent SDK for TypeScript - cost budgets, tamper-evident audit, and PII redaction built into the loop.
TypeScript★ 0↓ 1684/mes29 ago 2026
Apache-2.029 ago 2026 · métricas 2.10.0
PyPI
65Buenoíndice de salud
dunetrace/dunetrace
Real-time monitoring of production AI agents.
Python★ 5716 jul 2026
Licencia propia16 jul 2026 · métricas 2.10.0
Go
63Moderadoíndice de salud
fgn/go-langfuse
Langfuse client for Go, built on OpenTelemetry
Go★ 022 jul 2026
Apache-2.022 jul 2026 · métricas 2.10.0
npm
62Moderadoíndice de salud
Medhovarsh/forkmind
🧠 Local-first LLM state branching & debugging. Capture, visualize, branch, and regression-test LLM calls as a DAG. Free & local via Ollama, any OpenAI-compatible API, MCP for agents. Zero config.
JavaScript★ 2↓ 2178/mes24 jul 2026
MIT24 jul 2026 · métricas 2.10.0
npm
59Moderadoíndice de salud
TaewoooPark/Agent-Blackbox
Local-first flight recorder for coding agents : replay every run as a live session map, score the context bill, and write the fix back into AGENTS.md — no API key, one npx command.
TypeScript★ 57↓ 3988/mes21 jul 2026
MIT21 jul 2026 · métricas 2.10.0
Go · npm
56Moderadoíndice de salud
marmutapp/superbased-observer
Local-first cost & token tracking for Claude Code, Cursor, Codex & 23 more AI coding agents — proxy-accurate per-model spend, an MCP server your agent can query, and an opt-in team rollup. 100% local, no telemetry.
Go · TypeScript★ 3119 jul 2026
Licencia propia19 jul 2026 · métricas 2.10.0
npm
54Moderadoíndice de salud
inferock/inferock-bench
Local LLM cost-tracking proxy for OpenAI, Anthropic, Gemini, and pinned OpenRouter calls with token usage, failure, and billing-integrity receipts.
TypeScript★ 123↓ 6119/mes29 jul 2026
Licencia propia29 jul 2026 · métricas 2.10.0
npm
50Moderadoíndice de salud
PunkTechnologies/punk-sdk
Public TypeScript SDK and integration examples for Punk, the adaptive runtime for production AI agents.
TypeScript★ 1↓ 4631/mes3 sept 2026
Apache-2.03 sept 2026 · métricas 2.10.0