Усі теги
Тег каталогу

#asr

Усі репозиторії публічного реєстру з цим тегом — із тем GitHub або ключових слів, які публікують їхні реєстри пакетів. Здоров'я вимірюється за тією ж версіонованою методологією, що й решта реєстру.

27 записів
З тегом «asr»Упорядковано за індексом здоров'я
PyPI
99Винятковийіндекс здоров'я
NVIDIA-NeMo/Speech
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
Python · Jupyter Notebook★ 18.1K12 серп. 2026 р.
Apache-2.012 серп. 2026 р. · метрики 2.10.0
npm
94Винятковийіндекс здоров'я
deepgram/deepgram-js-sdk
Official JavaScript SDK for Deepgram.
TypeScript★ 272↓ 3.1M/міс27 серп. 2026 р.
MIT27 серп. 2026 р. · метрики 2.10.0
PyPI · npm
94Винятковийіндекс здоров'я
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.6K5 серп. 2026 р.
MIT5 серп. 2026 р. · метрики 2.10.0
PyPI
93Винятковийіндекс здоров'я
deepgram/deepgram-python-sdk
Official Python SDK for Deepgram.
Python★ 45919 серп. 2026 р.
MIT19 серп. 2026 р. · метрики 2.10.0
92Відміннийіндекс здоров'я
FluidInference/FluidAudio
Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
Swift★ 2 46716 лип. 2026 р.
Apache-2.016 лип. 2026 р. · метрики 2.10.0
npm
90Відміннийіндекс здоров'я
AssemblyAI/assemblyai-node-sdk
The AssemblyAI JavaScript SDK provides an easy-to-use interface for interacting with the AssemblyAI API, which supports async and real-time transcription, audio intelligence models, as well as the latest LeMUR models.
TypeScript★ 75↓ 2.2M/міс6 вер. 2026 р.
MIT6 вер. 2026 р. · метрики 2.10.0
PyPI
87Відміннийіндекс здоров'я
mbailey/voicemode
Natural voice conversations with Claude Code
Python★ 1 308↓ 22.8K/міс4 серп. 2026 р.
MIT4 серп. 2026 р. · метрики 2.10.0
PyPI · npm
87Відміннийіндекс здоров'я
moonshine-ai/moonshine
Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces
C++ · C★ 10.9K↓ 49.9K/міс28 серп. 2026 р.
Власна ліцензія28 серп. 2026 р. · метрики 2.10.0
npm · crates.io
86Відміннийіндекс здоров'я
drakulavich/kesha-voice-kit
Give your tools a voice — speech to text and back, 25 languages, up to ~19× faster than Whisper. On your machine.
Rust · TypeScript★ 67↓ 1 029/міс28 лип. 2026 р.
MIT28 лип. 2026 р. · метрики 2.10.0
PyPI
84Відміннийіндекс здоров'я
cmusphinx/pocketsphinx
A small speech recognizer
C★ 4 33212 серп. 2026 р.
Власна ліцензія12 серп. 2026 р. · метрики 2.10.0
crates.io · Maven · npm +1
84Відміннийіндекс здоров'я
k2-fsa/sherpa-onnx
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
C++ · Python★ 14.4K27 серп. 2026 р.
Apache-2.027 серп. 2026 р. · метрики 2.10.0
PyPI
83Відміннийіндекс здоров'я
modelscope/FunClip
FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
Python★ 6 20131 серп. 2026 р.
MIT31 серп. 2026 р. · метрики 2.10.0
PyPI
81Відміннийіндекс здоров'я
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
Python★ 347↓ 244.7K/міс20 лип. 2026 р.
MIT20 лип. 2026 р. · метрики 2.10.0
PyPI
81Відміннийіндекс здоров'я
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.4K↓ 1.5M/міс5 серп. 2026 р.
BSD-2-Clause5 серп. 2026 р. · метрики 2.10.0
PyPI · npm
78Добрийіндекс здоров'я
bengizmo/voxint
Опис репозиторію не опубліковано.
Python★ 3↓ 2 239/міс21 серп. 2026 р.
Apache-2.021 серп. 2026 р. · метрики 2.10.0
npm
78Добрийіндекс здоров'я
speechmatics/speechmatics-js-sdk
Javascript and Typescript SDK for Speechmatics
TypeScript · HTML★ 57↓ 244.1K/міс26 лип. 2026 р.
MIT26 лип. 2026 р. · метрики 2.10.0
crates.io
73Добрийіндекс здоров'я
altunenes/parakeet-rs
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
Rust★ 378↓ 12.5K/міс29 лип. 2026 р.
MIT29 лип. 2026 р. · метрики 2.10.0
PyPI · crates.io
67Добрийіндекс здоров'я
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
C++ · C★ 197↓ 7 559/міс17 лип. 2026 р.
MIT17 лип. 2026 р. · метрики 2.10.0
PyPI
65Добрийіндекс здоров'я
jdepoix/youtube-transcript-api
This is a python API which allows you to get the transcript/subtitles for a given YouTube video. It also works for automatically generated subtitles and it does not require an API key nor a headless browser, like other selenium based solutions do!
Python★ 8 121↓ 28.7M/міс27 серп. 2026 р.
MIT27 серп. 2026 р. · метрики 2.10.0
npm · crates.io
62Помірнийіндекс здоров'я
snomiao/otoji
realtime speech ⇄ text, 音を字に — wire mic → STT → translate → speech as an on-device voice graph that spans your devices over WebRTC. Runs in the browser (transformers.js/ONNX/WebGPU); no API keys, nothing leaves the device by default.
TypeScript · Rust★ 2↓ 11.3K/міс5 вер. 2026 р.
MIT5 вер. 2026 р. · метрики 2.10.0
Go
59Помірнийіндекс здоров'я
VoiceBlender/voiceblender
A programmable voice platform: SIP and WebRTC call control, multi-party mixing, recording, TTS/STT, and pluggable AI agents (ElevenLabs, VAPI, Pipecat, Deepgram) — all driven through a REST API, webhooks, and a WebSocket event stream
Go★ 9720 лип. 2026 р.
MIT20 лип. 2026 р. · метрики 2.10.0
NuGet
57Помірнийіндекс здоров'я
umlx5h/LLPlayer
The media player for language learning, with dual subtitles, AI-generated subtitles, real-time translation, and more!
C#★ 3 9905 серп. 2026 р.
GPL-3.05 серп. 2026 р. · метрики 2.10.0
Maven
51Помірнийіндекс здоров'я
CrispStrobe/CrisperWeaver
On-device speech-to-text Flutter app powered by CrispASR (ggml / Whisper) — offline, multi-platform, AGPL-3.0.
Dart★ 4130 лип. 2026 р.
AGPL-3.030 лип. 2026 р. · метрики 2.10.0
npm
41Слабкийіндекс здоров'я
surajmandalcell/asrpro
AI powered desktop transcription app with real time speech recognition, file transcription, global hotkeys, and SRT subtitle export.
TypeScript · JavaScript★ 41 серп. 2026 р.
Без ліцензії1 серп. 2026 р. · метрики 2.10.0
PyPI · npm · Go +2
28У зоні ризикуіндекс здоров'я
alphacep/vosk-api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Jupyter Notebook★ 15K↓ 776.1K/міс12 серп. 2026 р.
Apache-2.012 серп. 2026 р. · метрики 2.10.0
PyPI
23У зоні ризикуіндекс здоров'я
wiseman/py-webrtcvad
Python interface to the WebRTC Voice Activity Detector
C · C++★ 2 492↓ 451.3K/міс21 лип. 2026 р.
Власна ліцензія21 лип. 2026 р. · метрики 2.10.0
PyPI
11Критичнийіндекс здоров'я
abhirooptalasila/AutoSub
A CLI script to generate subtitle files (SRT/VTT/TXT) for any video using either DeepSpeech or Coqui
Python★ 6514 серп. 2026 р.
MIT4 серп. 2026 р. · метрики 2.10.0