Todas las etiquetas
Etiqueta del catálogo

#speech-recognition

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

28 registros
Con la etiqueta «speech-recognition»Ordenado por índice de salud
PyPI
99Excepcionalíndice de salud
huggingface/transformers
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Python★ 163.3K↓ 179M/mes4 ago 2026
Apache-2.04 ago 2026 · métricas 2.10.0
PyPI
95Excepcionalíndice de salud
Blaizzy/mlx-audio
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
Python★ 7796↓ 582.2K/mes28 ago 2026
MIT28 ago 2026 · métricas 2.10.0
npm
94Excepcionalíndice de salud
deepgram/deepgram-js-sdk
Official JavaScript SDK for Deepgram.
TypeScript★ 272↓ 3.1M/mes27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
PyPI · npm
94Excepcionalíndice de salud
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.6K5 ago 2026
MIT5 ago 2026 · métricas 2.10.0
PyPI
93Excepcionalíndice de salud
deepgram/deepgram-python-sdk
Official Python SDK for Deepgram.
Python★ 45919 ago 2026
MIT19 ago 2026 · métricas 2.10.0
PyPI · npm
92Excelenteíndice de salud
Picovoice/porcupine
On-device wake word detection powered by deep learning
Python · TypeScript · Swift★ 4911↓ 773K/mes12 ago 2026
Apache-2.012 ago 2026 · métricas 2.10.0
npm · Maven
89Excelenteíndice de salud
mybigday/whisper.rn
React Native binding of whisper.cpp.
C++ · C · Metal★ 799↓ 43.6K/mes1 ago 2026
MIT1 ago 2026 · métricas 2.10.0
NuGet · npm
87Excelenteíndice de salud
sandrohanea/whisper.net
Whisper.net. Speech to text made simple using Whisper Models
C#★ 94131 ago 2026
MIT31 ago 2026 · métricas 2.10.0
PyPI
86Excelenteíndice de salud
Uberi/speech_recognition
Speech recognition module for Python, supporting several engines and APIs, online and offline.
Python★ 8987↓ 9.2M/mes27 ago 2026
BSD-3-Clause27 ago 2026 · métricas 2.10.0
PyPI
84Excelenteíndice de salud
cmusphinx/pocketsphinx
A small speech recognizer
C★ 433212 ago 2026
Licencia propia12 ago 2026 · métricas 2.10.0
PyPI
84Excelenteíndice de salud
lhotse-speech/lhotse
Tools for handling multimodal data in machine learning projects.
Python★ 1146↓ 1.4M/mes13 ago 2026
Apache-2.013 ago 2026 · métricas 2.10.0
npm · PyPI · crates.io
83Excelenteíndice de salud
decibri/decibri
Cross-platform audio capture, playback, and voice activity detection for Python, Rust, and Node.js powered by a single Rust core.
Rust · Python · JavaScript★ 23↓ 19.3K/mes23 jul 2026
Apache-2.023 jul 2026 · métricas 2.10.0
PyPI
83Excelenteíndice de salud
modelscope/FunClip
FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
Python★ 620131 ago 2026
MIT31 ago 2026 · métricas 2.10.0
PyPI
81Excelenteíndice de salud
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
Python★ 347↓ 244.7K/mes20 jul 2026
MIT20 jul 2026 · métricas 2.10.0
PyPI
81Excelenteíndice de salud
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.4K↓ 1.5M/mes5 ago 2026
BSD-2-Clause5 ago 2026 · métricas 2.10.0
crates.io
73Buenoíndice de salud
altunenes/parakeet-rs
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
Rust★ 378↓ 12.5K/mes29 jul 2026
MIT29 jul 2026 · métricas 2.10.0
Go
73Buenoíndice de salud
gojargo/jargo
A WebRTC-native, audio-first conversational-AI framework for Go.
Go★ 2617 jul 2026
BSD-2-Clause17 jul 2026 · métricas 2.10.0
crates.io · PyPI
69Buenoíndice de salud
crispstrobe/crispasr
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
C++ · Python · C★ 43516 jul 2026
MIT16 jul 2026 · métricas 2.10.0
npm
65Buenoíndice de salud
mybigday/whisper.node
An another Node binding of whisper.cpp to make same API with whisper.rn as much as possible.
C++ · JavaScript · C★ 8↓ 43.9K/mes25 jul 2026
MIT25 jul 2026 · métricas 2.10.0
PyPI
63Moderadoíndice de salud
SYSTRAN/faster-whisper
Faster Whisper transcription with CTranslate2
Python★ 24.7K5 ago 2026
MIT5 ago 2026 · métricas 2.10.0
PyPI
60Moderadoíndice de salud
lnxusr1/kenzy
Smart Ai Voice Assistant written in Python
Python★ 3↓ 4013/mes1 ago 2026
MIT1 ago 2026 · métricas 2.10.0
PyPI
59Moderadoíndice de salud
KoljaB/RealtimeSTT
A robust, efficient, low-latency speech-to-text library with advanced voice activity detection, wake word activation and instant transcription.
Python★ 10.1K13 ago 2026
MIT13 ago 2026 · métricas 2.10.0
Go
59Moderadoíndice de salud
VoiceBlender/voiceblender
A programmable voice platform: SIP and WebRTC call control, multi-party mixing, recording, TTS/STT, and pluggable AI agents (ElevenLabs, VAPI, Pipecat, Deepgram) — all driven through a REST API, webhooks, and a WebSocket event stream
Go★ 9720 jul 2026
MIT20 jul 2026 · métricas 2.10.0
PyPI
38Débilíndice de salud
HenestrosaDev/audiotext
A desktop application that transcribes audio from files, microphone input or YouTube videos with the option to translate the content and create subtitles.
Python★ 35118 jul 2026
Licencia propia18 jul 2026 · métricas 2.10.0
npm
36Débilíndice de salud
TranscribeJs/transcribe.js
Monorepo for Transcribe.js
JavaScript · C++★ 5318 jul 2026
MIT18 jul 2026 · métricas 2.10.0
PyPI
34En riesgoíndice de salud
openvinotoolkit/openvino
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
C++★ 10.6K12 ago 2026
Apache-2.012 ago 2026 · métricas 2.10.0
PyPI
33En riesgoíndice de salud
collectivat/cmusphinx-models
Acoustic and language models for minorised languages.
Python · Shell★ 2622 jul 2026
AGPL-3.022 jul 2026 · métricas 2.10.0
PyPI · npm · Go +2
28En riesgoíndice de salud
alphacep/vosk-api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Jupyter Notebook★ 15K↓ 776.1K/mes12 ago 2026
Apache-2.012 ago 2026 · métricas 2.10.0