Todas las etiquetas
Etiqueta del catálogo

#asr

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

27 registros
Con la etiqueta «asr»Ordenado por índice de salud
PyPI
99Excepcionalíndice de salud
NVIDIA-NeMo/Speech
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
Python · Jupyter Notebook★ 18.1K12 ago 2026
Apache-2.012 ago 2026 · métricas 2.10.0
npm
94Excepcionalíndice de salud
deepgram/deepgram-js-sdk
Official JavaScript SDK for Deepgram.
TypeScript★ 272↓ 3.1M/mes27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
PyPI · npm
94Excepcionalíndice de salud
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.6K5 ago 2026
MIT5 ago 2026 · métricas 2.10.0
PyPI
93Excepcionalíndice de salud
deepgram/deepgram-python-sdk
Official Python SDK for Deepgram.
Python★ 45919 ago 2026
MIT19 ago 2026 · métricas 2.10.0
92Excelenteíndice de salud
FluidInference/FluidAudio
Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
Swift★ 246716 jul 2026
Apache-2.016 jul 2026 · métricas 2.10.0
npm
90Excelenteíndice de salud
AssemblyAI/assemblyai-node-sdk
The AssemblyAI JavaScript SDK provides an easy-to-use interface for interacting with the AssemblyAI API, which supports async and real-time transcription, audio intelligence models, as well as the latest LeMUR models.
TypeScript★ 75↓ 2.2M/mes6 sept 2026
MIT6 sept 2026 · métricas 2.10.0
PyPI
87Excelenteíndice de salud
mbailey/voicemode
Natural voice conversations with Claude Code
Python★ 1308↓ 22.8K/mes4 ago 2026
MIT4 ago 2026 · métricas 2.10.0
PyPI · npm
87Excelenteíndice de salud
moonshine-ai/moonshine
Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces
C++ · C★ 10.9K↓ 49.9K/mes28 ago 2026
Licencia propia28 ago 2026 · métricas 2.10.0
npm · crates.io
86Excelenteíndice de salud
drakulavich/kesha-voice-kit
Give your tools a voice — speech to text and back, 25 languages, up to ~19× faster than Whisper. On your machine.
Rust · TypeScript★ 67↓ 1029/mes28 jul 2026
MIT28 jul 2026 · métricas 2.10.0
PyPI
84Excelenteíndice de salud
cmusphinx/pocketsphinx
A small speech recognizer
C★ 433212 ago 2026
Licencia propia12 ago 2026 · métricas 2.10.0
crates.io · Maven · npm +1
84Excelenteíndice de salud
k2-fsa/sherpa-onnx
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
C++ · Python★ 14.4K27 ago 2026
Apache-2.027 ago 2026 · métricas 2.10.0
PyPI
83Excelenteíndice de salud
modelscope/FunClip
FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
Python★ 620131 ago 2026
MIT31 ago 2026 · métricas 2.10.0
PyPI
81Excelenteíndice de salud
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
Python★ 347↓ 244.7K/mes20 jul 2026
MIT20 jul 2026 · métricas 2.10.0
PyPI
81Excelenteíndice de salud
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.4K↓ 1.5M/mes5 ago 2026
BSD-2-Clause5 ago 2026 · métricas 2.10.0
PyPI · npm
78Buenoíndice de salud
bengizmo/voxint
El repositorio no publica descripción.
Python★ 3↓ 2239/mes21 ago 2026
Apache-2.021 ago 2026 · métricas 2.10.0
npm
78Buenoíndice de salud
speechmatics/speechmatics-js-sdk
Javascript and Typescript SDK for Speechmatics
TypeScript · HTML★ 57↓ 244.1K/mes26 jul 2026
MIT26 jul 2026 · métricas 2.10.0
crates.io
73Buenoíndice de salud
altunenes/parakeet-rs
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
Rust★ 378↓ 12.5K/mes29 jul 2026
MIT29 jul 2026 · métricas 2.10.0
PyPI · crates.io
67Buenoíndice de salud
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
C++ · C★ 197↓ 7559/mes17 jul 2026
MIT17 jul 2026 · métricas 2.10.0
PyPI
65Buenoíndice de salud
jdepoix/youtube-transcript-api
This is a python API which allows you to get the transcript/subtitles for a given YouTube video. It also works for automatically generated subtitles and it does not require an API key nor a headless browser, like other selenium based solutions do!
Python★ 8121↓ 28.7M/mes27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
npm · crates.io
62Moderadoíndice de salud
snomiao/otoji
realtime speech ⇄ text, 音を字に — wire mic → STT → translate → speech as an on-device voice graph that spans your devices over WebRTC. Runs in the browser (transformers.js/ONNX/WebGPU); no API keys, nothing leaves the device by default.
TypeScript · Rust★ 2↓ 11.3K/mes5 sept 2026
MIT5 sept 2026 · métricas 2.10.0
Go
59Moderadoíndice de salud
VoiceBlender/voiceblender
A programmable voice platform: SIP and WebRTC call control, multi-party mixing, recording, TTS/STT, and pluggable AI agents (ElevenLabs, VAPI, Pipecat, Deepgram) — all driven through a REST API, webhooks, and a WebSocket event stream
Go★ 9720 jul 2026
MIT20 jul 2026 · métricas 2.10.0
NuGet
57Moderadoíndice de salud
umlx5h/LLPlayer
The media player for language learning, with dual subtitles, AI-generated subtitles, real-time translation, and more!
C#★ 39905 ago 2026
GPL-3.05 ago 2026 · métricas 2.10.0
Maven
51Moderadoíndice de salud
CrispStrobe/CrisperWeaver
On-device speech-to-text Flutter app powered by CrispASR (ggml / Whisper) — offline, multi-platform, AGPL-3.0.
Dart★ 4130 jul 2026
AGPL-3.030 jul 2026 · métricas 2.10.0
npm
41Débilíndice de salud
surajmandalcell/asrpro
AI powered desktop transcription app with real time speech recognition, file transcription, global hotkeys, and SRT subtitle export.
TypeScript · JavaScript★ 41 ago 2026
Sin licencia1 ago 2026 · métricas 2.10.0
PyPI · npm · Go +2
28En riesgoíndice de salud
alphacep/vosk-api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Jupyter Notebook★ 15K↓ 776.1K/mes12 ago 2026
Apache-2.012 ago 2026 · métricas 2.10.0
PyPI
23En riesgoíndice de salud
wiseman/py-webrtcvad
Python interface to the WebRTC Voice Activity Detector
C · C++★ 2492↓ 451.3K/mes21 jul 2026
Licencia propia21 jul 2026 · métricas 2.10.0
PyPI
11Críticoíndice de salud
abhirooptalasila/AutoSub
A CLI script to generate subtitle files (SRT/VTT/TXT) for any video using either DeepSpeech or Coqui
Python★ 6514 ago 2026
MIT4 ago 2026 · métricas 2.10.0