Todas las etiquetas
Etiqueta del catálogo

#speech-recognition

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

15 registros
Con la etiqueta «speech-recognition»Ordenado por índice de salud
PyPI
88Excelenteíndice de salud
huggingface/transformers
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Python★ 162.7K20 jul 2026
Apache-2.020 jul 2026 · métricas 1.13.0
PyPI
87Excelenteíndice de salud
huggingface/pytorch-pretrained-BERT
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Python★ 162.7K16 jul 2026
Apache-2.016 jul 2026 · métricas 1.13.0
npm
79Buenoíndice de salud
deepgram/deepgram-js-sdk
Official JavaScript SDK for Deepgram.
TypeScript★ 268↓ 2.4M/mes19 jul 2026
MIT19 jul 2026 · métricas 1.13.0
PyPI · npm
78Buenoíndice de salud
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.4K21 jul 2026
MIT21 jul 2026 · métricas 1.13.0
npm · PyPI
77Buenoíndice de salud
alibaba-damo-academy/funasr
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.3K17 jul 2026
MIT17 jul 2026 · métricas 1.13.0
npm · PyPI
77Buenoíndice de salud
alibaba/funasr
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.3K17 jul 2026
MIT17 jul 2026 · métricas 1.13.0
PyPI
70Buenoíndice de salud
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.2K21 jul 2026
BSD-2-Clause21 jul 2026 · métricas 1.13.0
Go
63Moderadoíndice de salud
gojargo/jargo
A WebRTC-native, audio-first conversational-AI framework for Go.
Go★ 2617 jul 2026
BSD-2-Clause17 jul 2026 · métricas 1.13.0
crates.io · PyPI
61Moderadoíndice de salud
crispstrobe/crispasr
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
C++ · Python · C★ 43516 jul 2026
MIT16 jul 2026 · métricas 1.13.0
PyPI
57Moderadoíndice de salud
SYSTRAN/faster-whisper
Faster Whisper transcription with CTranslate2
Python★ 24.4K18 jul 2026
MIT18 jul 2026 · métricas 1.13.0
Go
55Moderadoíndice de salud
VoiceBlender/voiceblender
A programmable voice platform: SIP and WebRTC call control, multi-party mixing, recording, TTS/STT, and pluggable AI agents (ElevenLabs, VAPI, Pipecat, Deepgram) — all driven through a REST API, webhooks, and a WebSocket event stream
Go★ 9720 jul 2026
MIT20 jul 2026 · métricas 1.13.0
PyPI
44En riesgoíndice de salud
HenestrosaDev/audiotext
A desktop application that transcribes audio from files, microphone input or YouTube videos with the option to translate the content and create subtitles.
Python★ 35118 jul 2026
Licencia propia18 jul 2026 · métricas 1.13.0
npm
41En riesgoíndice de salud
TranscribeJs/transcribe.js
Monorepo for Transcribe.js
JavaScript · C++★ 5318 jul 2026
MIT18 jul 2026 · métricas 1.13.0
PyPI
39En riesgoíndice de salud
collectivat/cmusphinx-models
Acoustic and language models for minorised languages.
Python · Shell★ 2622 jul 2026
AGPL-3.022 jul 2026 · métricas 1.13.0
PyPI
33En riesgoíndice de salud
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
Python★ 347↓ 244.7K/mes20 jul 2026
MIT20 jul 2026 · métricas 1.13.0