Todas las etiquetas
Etiqueta del catálogo

#transcription

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

32 registros
Con la etiqueta «transcription»Ordenado por índice de salud
npm
95Excepcionalíndice de salud
TanStack/ai
🤖 Type-safe, provider-agnostic TypeScript AI SDK for streaming chat, tool calling, agents, and multimodal apps across OpenAI, Anthropic, Gemini, React, Vue, Svelte, and Solid.
TypeScript★ 2907↓ 613.6K/mes22 jul 2026
MIT22 jul 2026 · métricas 2.10.0
PyPI · npm
94Excepcionalíndice de salud
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.6K5 ago 2026
MIT5 ago 2026 · métricas 2.10.0
npm
90Excelenteíndice de salud
AssemblyAI/assemblyai-node-sdk
The AssemblyAI JavaScript SDK provides an easy-to-use interface for interacting with the AssemblyAI API, which supports async and real-time transcription, audio intelligence models, as well as the latest LeMUR models.
TypeScript★ 75↓ 2.2M/mes6 sept 2026
MIT6 sept 2026 · métricas 2.10.0
PyPI · npm
87Excelenteíndice de salud
moonshine-ai/moonshine
Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces
C++ · C★ 10.9K↓ 49.9K/mes28 ago 2026
Licencia propia28 ago 2026 · métricas 2.10.0
npm · PyPI
86Excelenteíndice de salud
KyaniteLabs/kinocut
Guardrailed video editing MCP server for AI agents. FFmpeg, Hyperframes, repurposing tools, Python client, and CLI. Local, fast, free.
Python★ 88↓ 630/mes24 jul 2026
Apache-2.024 jul 2026 · métricas 2.10.0
npm · crates.io
86Excelenteíndice de salud
drakulavich/kesha-voice-kit
Give your tools a voice — speech to text and back, 25 languages, up to ~19× faster than Whisper. On your machine.
Rust · TypeScript★ 67↓ 1029/mes28 jul 2026
MIT28 jul 2026 · métricas 2.10.0
PyPI
83Excelenteíndice de salud
modelscope/FunClip
FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
Python★ 620131 ago 2026
MIT31 ago 2026 · métricas 2.10.0
npm · crates.io
83Excelenteíndice de salud
silverstein/minutes
Every meeting, every idea, every voice note — searchable by your AI. Open-source, privacy-first conversation memory layer.
Rust★ 1372↓ 5342/mes23 jul 2026
MIT23 jul 2026 · métricas 2.10.0
PyPI · npm
78Buenoíndice de salud
bengizmo/voxint
El repositorio no publica descripción.
Python★ 3↓ 2239/mes21 ago 2026
Apache-2.021 ago 2026 · métricas 2.10.0
npm
78Buenoíndice de salud
speechmatics/speechmatics-js-sdk
Javascript and Typescript SDK for Speechmatics
TypeScript · HTML★ 57↓ 244.1K/mes26 jul 2026
MIT26 jul 2026 · métricas 2.10.0
npm
78Buenoíndice de salud
zerodytrash/TikTok-Live-Connector
Node.js library to receive live stream events (comments, gifts, etc.) in realtime from TikTok LIVE.
TypeScript★ 2080↓ 110.9K/mes21 jul 2026
AGPL-3.021 jul 2026 · métricas 2.10.0
NuGet
73Buenoíndice de salud
GeiserX/whisper-subs
Jellyfin plugin for local AI-powered subtitle generation using Whisper - all processing stays on your server
C# · HTML★ 8124 jul 2026
GPL-3.024 jul 2026 · métricas 2.10.0
npm
73Buenoíndice de salud
guimatheus92/mcp-video-analyzer
MCP server that turns any video — YouTube, Instagram, TikTok, Loom, X, Vimeo, direct URLs, local files — into transcripts, key frames, OCR text, and metadata for AI agents.
TypeScript★ 28↓ 2780/mes27 jul 2026
MIT27 jul 2026 · métricas 2.10.0
npm · PyPI
71Buenoíndice de salud
resolvicomai/kassinao
Open-source Discord bot for named transcripts, meeting notes, tasks, and sourced answers.
TypeScript · Shell★ 1↓ 2119/mes20 jul 2026
AGPL-3.020 jul 2026 · métricas 2.10.0
crates.io · PyPI
69Buenoíndice de salud
crispstrobe/crispasr
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
C++ · Python · C★ 43516 jul 2026
MIT16 jul 2026 · métricas 2.10.0
npm · RubyGems
69Buenoíndice de salud
open-data-rescue/climate-data-rescue
Climate Data Rescue is an archival data rescue platform using Ruby on Rails.
JavaScript · Ruby · HTML★ 1528 jul 2026
MIT28 jul 2026 · métricas 2.10.0
PyPI · crates.io
67Buenoíndice de salud
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
C++ · C★ 197↓ 7559/mes17 jul 2026
MIT17 jul 2026 · métricas 2.10.0
npm
65Buenoíndice de salud
hexinmiao96/vue-ai-hooks
El repositorio no publica descripción.
TypeScript · JavaScript★ 1↓ 3121/mes31 jul 2026
MIT31 jul 2026 · métricas 2.10.0
npm
65Buenoíndice de salud
mybigday/whisper.node
An another Node binding of whisper.cpp to make same API with whisper.rn as much as possible.
C++ · JavaScript · C★ 8↓ 43.9K/mes25 jul 2026
MIT25 jul 2026 · métricas 2.10.0
npm · PyPI
63Moderadoíndice de salud
arcforgelabs/dictate
Desktop dictation that types into the focused app. Configurable push-to-talk, local/API transcription, tray settings, and recent history.
Python · JavaScript★ 0↓ 8233/mes16 jul 2026
MIT16 jul 2026 · métricas 2.10.0
PyPI · crates.io
63Moderadoíndice de salud
solpbc/solstone-journal
Navigate Life Intelligently
Python★ 17↓ 6851/mes1 ago 2026
AGPL-3.01 ago 2026 · métricas 2.10.0
PyPI
62Moderadoíndice de salud
HUANGCHIHHUNGLeo/claude-real-video
Let Claude (or any LLM) actually watch a video — scene-aware, deduplicated frames + transcript, from a URL or local file. Runs locally, MIT.
Python★ 2095↓ 6592/mes31 ago 2026
MIT31 ago 2026 · métricas 2.10.0
npm · crates.io
62Moderadoíndice de salud
snomiao/otoji
realtime speech ⇄ text, 音を字に — wire mic → STT → translate → speech as an on-device voice graph that spans your devices over WebRTC. Runs in the browser (transformers.js/ONNX/WebGPU); no API keys, nothing leaves the device by default.
TypeScript · Rust★ 2↓ 11.3K/mes5 sept 2026
MIT5 sept 2026 · métricas 2.10.0
PyPI
59Moderadoíndice de salud
KoljaB/RealtimeSTT
A robust, efficient, low-latency speech-to-text library with advanced voice activity detection, wake word activation and instant transcription.
Python★ 10.1K13 ago 2026
MIT13 ago 2026 · métricas 2.10.0
npm
53Moderadoíndice de salud
arach/vox
Local-first macOS transcription runtime with Swift services, a Bun CLI, and a TypeScript SDK.
Swift · TypeScript★ 519 jul 2026
Sin licencia19 jul 2026 · métricas 2.10.0
npm
53Moderadoíndice de salud
speakai/speakai-mcp
Speak MCP Server
TypeScript · HTML★ 0↓ 3146/mes19 jul 2026
Sin licencia19 jul 2026 · métricas 2.10.0
Maven
51Moderadoíndice de salud
CrispStrobe/CrisperWeaver
On-device speech-to-text Flutter app powered by CrispASR (ggml / Whisper) — offline, multi-platform, AGPL-3.0.
Dart★ 4130 jul 2026
AGPL-3.030 jul 2026 · métricas 2.10.0
Go
51Moderadoíndice de salud
emiliopalmerini/podscribe
podscribe is a small Go CLI for transcribing podcast audio with the ElevenLabs Speech to Text API.
Go★ 028 jul 2026
MIT28 jul 2026 · métricas 2.10.0
Go
48Débilíndice de salud
nerveband/vflow
Agent-native Go CLI for local-first video editing: canonical JSON artifacts, safe preview rendering, transcript/framing tools, provider QA, and NLE exchange.
Go★ 05 sept 2026
Sin licencia5 sept 2026 · métricas 2.10.0