全部标签
目录标签

#speech-recognition

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

28 条记录
标签为“speech-recognition”按健康指数排序
PyPI
99卓越健康指数
huggingface/transformers
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Python★ 163.3K↓ 179M/月2026年8月4日
Apache-2.02026年8月4日 · 指标 2.10.0
PyPI
95卓越健康指数
Blaizzy/mlx-audio
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
Python★ 7,796↓ 582.2K/月2026年8月28日
MIT2026年8月28日 · 指标 2.10.0
npm
94卓越健康指数
deepgram/deepgram-js-sdk
Official JavaScript SDK for Deepgram.
TypeScript★ 272↓ 3.1M/月2026年8月27日
MIT2026年8月27日 · 指标 2.10.0
PyPI · npm
94卓越健康指数
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.6K2026年8月5日
MIT2026年8月5日 · 指标 2.10.0
PyPI
93卓越健康指数
deepgram/deepgram-python-sdk
Official Python SDK for Deepgram.
Python★ 4592026年8月19日
MIT2026年8月19日 · 指标 2.10.0
PyPI · npm
92优秀健康指数
Picovoice/porcupine
On-device wake word detection powered by deep learning
Python · TypeScript · Swift★ 4,911↓ 773K/月2026年8月12日
Apache-2.02026年8月12日 · 指标 2.10.0
npm · Maven
89优秀健康指数
mybigday/whisper.rn
React Native binding of whisper.cpp.
C++ · C · Metal★ 799↓ 43.6K/月2026年8月1日
MIT2026年8月1日 · 指标 2.10.0
NuGet · npm
87优秀健康指数
sandrohanea/whisper.net
Whisper.net. Speech to text made simple using Whisper Models
C#★ 9412026年8月31日
MIT2026年8月31日 · 指标 2.10.0
PyPI
86优秀健康指数
Uberi/speech_recognition
Speech recognition module for Python, supporting several engines and APIs, online and offline.
Python★ 8,987↓ 9.2M/月2026年8月27日
BSD-3-Clause2026年8月27日 · 指标 2.10.0
PyPI
84优秀健康指数
cmusphinx/pocketsphinx
A small speech recognizer
C★ 4,3322026年8月12日
自定义许可证2026年8月12日 · 指标 2.10.0
PyPI
84优秀健康指数
lhotse-speech/lhotse
Tools for handling multimodal data in machine learning projects.
Python★ 1,146↓ 1.4M/月2026年8月13日
Apache-2.02026年8月13日 · 指标 2.10.0
npm · PyPI · crates.io
83优秀健康指数
decibri/decibri
Cross-platform audio capture, playback, and voice activity detection for Python, Rust, and Node.js powered by a single Rust core.
Rust · Python · JavaScript★ 23↓ 19.3K/月2026年7月23日
Apache-2.02026年7月23日 · 指标 2.10.0
PyPI
83优秀健康指数
modelscope/FunClip
FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
Python★ 6,2012026年8月31日
MIT2026年8月31日 · 指标 2.10.0
PyPI
81优秀健康指数
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
Python★ 347↓ 244.7K/月2026年7月20日
MIT2026年7月20日 · 指标 2.10.0
PyPI
81优秀健康指数
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.4K↓ 1.5M/月2026年8月5日
BSD-2-Clause2026年8月5日 · 指标 2.10.0
crates.io
73良好健康指数
altunenes/parakeet-rs
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
Rust★ 378↓ 12.5K/月2026年7月29日
MIT2026年7月29日 · 指标 2.10.0
Go
73良好健康指数
gojargo/jargo
A WebRTC-native, audio-first conversational-AI framework for Go.
Go★ 262026年7月17日
BSD-2-Clause2026年7月17日 · 指标 2.10.0
crates.io · PyPI
69良好健康指数
crispstrobe/crispasr
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
C++ · Python · C★ 4352026年7月16日
MIT2026年7月16日 · 指标 2.10.0
npm
65良好健康指数
mybigday/whisper.node
An another Node binding of whisper.cpp to make same API with whisper.rn as much as possible.
C++ · JavaScript · C★ 8↓ 43.9K/月2026年7月25日
MIT2026年7月25日 · 指标 2.10.0
PyPI
63中等健康指数
SYSTRAN/faster-whisper
Faster Whisper transcription with CTranslate2
Python★ 24.7K2026年8月5日
MIT2026年8月5日 · 指标 2.10.0
PyPI
60中等健康指数
lnxusr1/kenzy
Smart Ai Voice Assistant written in Python
Python★ 3↓ 4,013/月2026年8月1日
MIT2026年8月1日 · 指标 2.10.0
PyPI
59中等健康指数
KoljaB/RealtimeSTT
A robust, efficient, low-latency speech-to-text library with advanced voice activity detection, wake word activation and instant transcription.
Python★ 10.1K2026年8月13日
MIT2026年8月13日 · 指标 2.10.0
Go
59中等健康指数
VoiceBlender/voiceblender
A programmable voice platform: SIP and WebRTC call control, multi-party mixing, recording, TTS/STT, and pluggable AI agents (ElevenLabs, VAPI, Pipecat, Deepgram) — all driven through a REST API, webhooks, and a WebSocket event stream
Go★ 972026年7月20日
MIT2026年7月20日 · 指标 2.10.0
PyPI
38薄弱健康指数
HenestrosaDev/audiotext
A desktop application that transcribes audio from files, microphone input or YouTube videos with the option to translate the content and create subtitles.
Python★ 3512026年7月18日
自定义许可证2026年7月18日 · 指标 2.10.0
npm
36薄弱健康指数
TranscribeJs/transcribe.js
Monorepo for Transcribe.js
JavaScript · C++★ 532026年7月18日
MIT2026年7月18日 · 指标 2.10.0
PyPI
34存在风险健康指数
openvinotoolkit/openvino
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
C++★ 10.6K2026年8月12日
Apache-2.02026年8月12日 · 指标 2.10.0
PyPI
33存在风险健康指数
collectivat/cmusphinx-models
Acoustic and language models for minorised languages.
Python · Shell★ 262026年7月22日
AGPL-3.02026年7月22日 · 指标 2.10.0
PyPI · npm · Go +2
28存在风险健康指数
alphacep/vosk-api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Jupyter Notebook★ 15K↓ 776.1K/月2026年8月12日
Apache-2.02026年8月12日 · 指标 2.10.0