All tags
Catalogue tag

#asr

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

27 records
Tagged “asr”Ranked by health index
PyPI
99Exceptionalhealth index
NVIDIA-NeMo/Speech
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
Python · Jupyter Notebook★ 18.1KAug 12, 2026
Apache-2.0Aug 12, 2026 · metrics 2.10.0
npm
94Exceptionalhealth index
deepgram/deepgram-js-sdk
Official JavaScript SDK for Deepgram.
TypeScript★ 272↓ 3.1M/moAug 27, 2026
MITAug 27, 2026 · metrics 2.10.0
PyPI · npm
94Exceptionalhealth index
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.6KAug 5, 2026
MITAug 5, 2026 · metrics 2.10.0
PyPI
93Exceptionalhealth index
deepgram/deepgram-python-sdk
Official Python SDK for Deepgram.
Python★ 459Aug 19, 2026
MITAug 19, 2026 · metrics 2.10.0
92Excellenthealth index
FluidInference/FluidAudio
Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
Swift★ 2,467Jul 16, 2026
Apache-2.0Jul 16, 2026 · metrics 2.10.0
npm
90Excellenthealth index
AssemblyAI/assemblyai-node-sdk
The AssemblyAI JavaScript SDK provides an easy-to-use interface for interacting with the AssemblyAI API, which supports async and real-time transcription, audio intelligence models, as well as the latest LeMUR models.
TypeScript★ 75↓ 2.2M/moSep 6, 2026
MITSep 6, 2026 · metrics 2.10.0
PyPI
87Excellenthealth index
mbailey/voicemode
Natural voice conversations with Claude Code
Python★ 1,308↓ 22.8K/moAug 4, 2026
MITAug 4, 2026 · metrics 2.10.0
PyPI · npm
87Excellenthealth index
moonshine-ai/moonshine
Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces
C++ · C★ 10.9K↓ 49.9K/moAug 28, 2026
Custom licenseAug 28, 2026 · metrics 2.10.0
npm · crates.io
86Excellenthealth index
drakulavich/kesha-voice-kit
Give your tools a voice — speech to text and back, 25 languages, up to ~19× faster than Whisper. On your machine.
Rust · TypeScript★ 67↓ 1,029/moJul 28, 2026
MITJul 28, 2026 · metrics 2.10.0
PyPI
84Excellenthealth index
cmusphinx/pocketsphinx
A small speech recognizer
C★ 4,332Aug 12, 2026
Custom licenseAug 12, 2026 · metrics 2.10.0
crates.io · Maven · npm +1
84Excellenthealth index
k2-fsa/sherpa-onnx
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
C++ · Python★ 14.4KAug 27, 2026
Apache-2.0Aug 27, 2026 · metrics 2.10.0
PyPI
83Excellenthealth index
modelscope/FunClip
FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
Python★ 6,201Aug 31, 2026
MITAug 31, 2026 · metrics 2.10.0
PyPI
81Excellenthealth index
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
Python★ 347↓ 244.7K/moJul 20, 2026
MITJul 20, 2026 · metrics 2.10.0
PyPI
81Excellenthealth index
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.4K↓ 1.5M/moAug 5, 2026
BSD-2-ClauseAug 5, 2026 · metrics 2.10.0
PyPI · npm
78Goodhealth index
bengizmo/voxint
No repository description published.
Python★ 3↓ 2,239/moAug 21, 2026
Apache-2.0Aug 21, 2026 · metrics 2.10.0
npm
78Goodhealth index
speechmatics/speechmatics-js-sdk
Javascript and Typescript SDK for Speechmatics
TypeScript · HTML★ 57↓ 244.1K/moJul 26, 2026
MITJul 26, 2026 · metrics 2.10.0
crates.io
73Goodhealth index
altunenes/parakeet-rs
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
Rust★ 378↓ 12.5K/moJul 29, 2026
MITJul 29, 2026 · metrics 2.10.0
PyPI · crates.io
67Goodhealth index
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
C++ · C★ 197↓ 7,559/moJul 17, 2026
MITJul 17, 2026 · metrics 2.10.0
PyPI
65Goodhealth index
jdepoix/youtube-transcript-api
This is a python API which allows you to get the transcript/subtitles for a given YouTube video. It also works for automatically generated subtitles and it does not require an API key nor a headless browser, like other selenium based solutions do!
Python★ 8,121↓ 28.7M/moAug 27, 2026
MITAug 27, 2026 · metrics 2.10.0
npm · crates.io
62Moderatehealth index
snomiao/otoji
realtime speech ⇄ text, 音を字に — wire mic → STT → translate → speech as an on-device voice graph that spans your devices over WebRTC. Runs in the browser (transformers.js/ONNX/WebGPU); no API keys, nothing leaves the device by default.
TypeScript · Rust★ 2↓ 11.3K/moSep 5, 2026
MITSep 5, 2026 · metrics 2.10.0
Go
59Moderatehealth index
VoiceBlender/voiceblender
A programmable voice platform: SIP and WebRTC call control, multi-party mixing, recording, TTS/STT, and pluggable AI agents (ElevenLabs, VAPI, Pipecat, Deepgram) — all driven through a REST API, webhooks, and a WebSocket event stream
Go★ 97Jul 20, 2026
MITJul 20, 2026 · metrics 2.10.0
NuGet
57Moderatehealth index
umlx5h/LLPlayer
The media player for language learning, with dual subtitles, AI-generated subtitles, real-time translation, and more!
C#★ 3,990Aug 5, 2026
GPL-3.0Aug 5, 2026 · metrics 2.10.0
Maven
51Moderatehealth index
CrispStrobe/CrisperWeaver
On-device speech-to-text Flutter app powered by CrispASR (ggml / Whisper) — offline, multi-platform, AGPL-3.0.
Dart★ 41Jul 30, 2026
AGPL-3.0Jul 30, 2026 · metrics 2.10.0
npm
41Weakhealth index
surajmandalcell/asrpro
AI powered desktop transcription app with real time speech recognition, file transcription, global hotkeys, and SRT subtitle export.
TypeScript · JavaScript★ 4Aug 1, 2026
No licenseAug 1, 2026 · metrics 2.10.0
PyPI · npm · Go +2
28At Riskhealth index
alphacep/vosk-api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Jupyter Notebook★ 15K↓ 776.1K/moAug 12, 2026
Apache-2.0Aug 12, 2026 · metrics 2.10.0
PyPI
23At Riskhealth index
wiseman/py-webrtcvad
Python interface to the WebRTC Voice Activity Detector
C · C++★ 2,492↓ 451.3K/moJul 21, 2026
Custom licenseJul 21, 2026 · metrics 2.10.0
PyPI
11Criticalhealth index
abhirooptalasila/AutoSub
A CLI script to generate subtitle files (SRT/VTT/TXT) for any video using either DeepSpeech or Coqui
Python★ 651Aug 4, 2026
MITAug 4, 2026 · metrics 2.10.0