全部标签
目录标签

#asr

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

14 条记录
标签为“asr”按健康指数排序
npm
79良好健康指数
deepgram/deepgram-js-sdk
Official JavaScript SDK for Deepgram.
TypeScript★ 268↓ 2.4M/月2026年7月19日
MIT2026年7月19日 · 指标 1.13.0
78良好健康指数
FluidInference/FluidAudio
Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
Swift★ 2,4672026年7月16日
Apache-2.02026年7月16日 · 指标 1.13.0
PyPI · npm
78良好健康指数
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.4K2026年7月21日
MIT2026年7月21日 · 指标 1.13.0
npm · PyPI
77良好健康指数
alibaba-damo-academy/funasr
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.3K2026年7月17日
MIT2026年7月17日 · 指标 1.13.0
npm · PyPI
77良好健康指数
alibaba/funasr
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.3K2026年7月17日
MIT2026年7月17日 · 指标 1.13.0
npm
73良好健康指数
AssemblyAI/assemblyai-node-sdk
The AssemblyAI JavaScript SDK provides an easy-to-use interface for interacting with the AssemblyAI API, which supports async and real-time transcription, audio intelligence models, as well as the latest LeMUR models.
TypeScript★ 76↓ 1.6M/月2026年7月15日
MIT2026年7月15日 · 指标 1.13.0
PyPI
70良好健康指数
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.2K2026年7月21日
BSD-2-Clause2026年7月21日 · 指标 1.13.0
crates.io · Maven · npm +1
69中等健康指数
k2-fsa/sherpa-onnx
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
C++ · Python★ 13.6K2026年7月17日
Apache-2.02026年7月17日 · 指标 1.13.0
PyPI
65中等健康指数
jdepoix/youtube-transcript-api
This is a python API which allows you to get the transcript/subtitles for a given YouTube video. It also works for automatically generated subtitles and it does not require an API key nor a headless browser, like other selenium based solutions do!
Python★ 7,938↓ 32.2M/月2026年7月21日
MIT2026年7月21日 · 指标 1.13.0
PyPI · crates.io
61中等健康指数
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
C++ · C★ 197↓ 7,559/月2026年7月17日
MIT2026年7月17日 · 指标 1.13.0
npm · crates.io
59中等健康指数
snomiao/otoji
realtime speech ⇄ text, 音を字に — wire mic → STT → translate → speech as an on-device voice graph that spans your devices over WebRTC. Runs in the browser (transformers.js/ONNX/WebGPU); no API keys, nothing leaves the device by default.
TypeScript · Rust★ 1↓ 4,782/月2026年7月15日
MIT2026年7月15日 · 指标 1.13.0
Go
55中等健康指数
VoiceBlender/voiceblender
A programmable voice platform: SIP and WebRTC call control, multi-party mixing, recording, TTS/STT, and pluggable AI agents (ElevenLabs, VAPI, Pipecat, Deepgram) — all driven through a REST API, webhooks, and a WebSocket event stream
Go★ 972026年7月20日
MIT2026年7月20日 · 指标 1.13.0
PyPI
33存在风险健康指数
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
Python★ 347↓ 244.7K/月2026年7月20日
MIT2026年7月20日 · 指标 1.13.0
PyPI
20危急健康指数
wiseman/py-webrtcvad
Python interface to the WebRTC Voice Activity Detector
C · C++★ 2,492↓ 451.3K/月2026年7月21日
自定义许可证2026年7月21日 · 指标 1.13.0