全部标签
目录标签

#speech-recognition

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

16 条记录
标签为“speech-recognition”按健康指数排序
PyPI
88优秀健康指数
huggingface/transformers
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Python★ 162.7K2026年7月20日
Apache-2.02026年7月20日 · 指标 1.13.0
PyPI
87优秀健康指数
huggingface/pytorch-pretrained-BERT
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Python★ 162.7K2026年7月16日
Apache-2.02026年7月16日 · 指标 1.13.0
npm
79良好健康指数
deepgram/deepgram-js-sdk
Official JavaScript SDK for Deepgram.
TypeScript★ 268↓ 2.4M/月2026年7月19日
MIT2026年7月19日 · 指标 1.13.0
PyPI · npm
78良好健康指数
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.4K2026年7月21日
MIT2026年7月21日 · 指标 1.13.0
npm · PyPI
77良好健康指数
alibaba-damo-academy/funasr
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.3K2026年7月17日
MIT2026年7月17日 · 指标 1.13.0
npm · PyPI
77良好健康指数
alibaba/funasr
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.3K2026年7月17日
MIT2026年7月17日 · 指标 1.13.0
PyPI
70良好健康指数
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.2K2026年7月21日
BSD-2-Clause2026年7月21日 · 指标 1.13.0
npm · PyPI · crates.io
69中等健康指数
decibri/decibri
Cross-platform audio capture, playback, and voice activity detection for Python, Rust, and Node.js powered by a single Rust core.
Rust · Python · JavaScript★ 23↓ 19.3K/月2026年7月23日
Apache-2.02026年7月23日 · 指标 1.13.0
Go
63中等健康指数
gojargo/jargo
A WebRTC-native, audio-first conversational-AI framework for Go.
Go★ 262026年7月17日
BSD-2-Clause2026年7月17日 · 指标 1.13.0
crates.io · PyPI
61中等健康指数
crispstrobe/crispasr
C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more
C++ · Python · C★ 4352026年7月16日
MIT2026年7月16日 · 指标 1.13.0
PyPI
57中等健康指数
SYSTRAN/faster-whisper
Faster Whisper transcription with CTranslate2
Python★ 24.4K2026年7月18日
MIT2026年7月18日 · 指标 1.13.0
Go
55中等健康指数
VoiceBlender/voiceblender
A programmable voice platform: SIP and WebRTC call control, multi-party mixing, recording, TTS/STT, and pluggable AI agents (ElevenLabs, VAPI, Pipecat, Deepgram) — all driven through a REST API, webhooks, and a WebSocket event stream
Go★ 972026年7月20日
MIT2026年7月20日 · 指标 1.13.0
PyPI
44存在风险健康指数
HenestrosaDev/audiotext
A desktop application that transcribes audio from files, microphone input or YouTube videos with the option to translate the content and create subtitles.
Python★ 3512026年7月18日
自定义许可证2026年7月18日 · 指标 1.13.0
npm
41存在风险健康指数
TranscribeJs/transcribe.js
Monorepo for Transcribe.js
JavaScript · C++★ 532026年7月18日
MIT2026年7月18日 · 指标 1.13.0
PyPI
39存在风险健康指数
collectivat/cmusphinx-models
Acoustic and language models for minorised languages.
Python · Shell★ 262026年7月22日
AGPL-3.02026年7月22日 · 指标 1.13.0
PyPI
33存在风险健康指数
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
Python★ 347↓ 244.7K/月2026年7月20日
MIT2026年7月20日 · 指标 1.13.0