Alle Tags
Katalog-Tag

#asr

Alle Repositories im öffentlichen Register, die dieses Tag tragen — aus ihren GitHub-Topics oder den von ihren Paket-Registries veröffentlichten Schlagwörtern. Die Gesundheit wird nach derselben versionierten Methodik gemessen wie im übrigen Register.

27 Einträge
Getaggt als „asr“Geordnet nach Gesundheitsindex
PyPI
99AußergewöhnlichGesundheitsindex
NVIDIA-NeMo/Speech
A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)
Python · Jupyter Notebook★ 18.1K12. Aug. 2026
Apache-2.012. Aug. 2026 · Metriken 2.10.0
npm
94AußergewöhnlichGesundheitsindex
deepgram/deepgram-js-sdk
Official JavaScript SDK for Deepgram.
TypeScript★ 272↓ 3.1M/Monat27. Aug. 2026
MIT27. Aug. 2026 · Metriken 2.10.0
PyPI · npm
94AußergewöhnlichGesundheitsindex
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.6K5. Aug. 2026
MIT5. Aug. 2026 · Metriken 2.10.0
PyPI
93AußergewöhnlichGesundheitsindex
deepgram/deepgram-python-sdk
Official Python SDK for Deepgram.
Python★ 45919. Aug. 2026
MIT19. Aug. 2026 · Metriken 2.10.0
92ExzellentGesundheitsindex
FluidInference/FluidAudio
Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
Swift★ 2.46716. Juli 2026
Apache-2.016. Juli 2026 · Metriken 2.10.0
npm
90ExzellentGesundheitsindex
AssemblyAI/assemblyai-node-sdk
The AssemblyAI JavaScript SDK provides an easy-to-use interface for interacting with the AssemblyAI API, which supports async and real-time transcription, audio intelligence models, as well as the latest LeMUR models.
TypeScript★ 75↓ 2.2M/Monat6. Sept. 2026
MIT6. Sept. 2026 · Metriken 2.10.0
PyPI
87ExzellentGesundheitsindex
mbailey/voicemode
Natural voice conversations with Claude Code
Python★ 1.308↓ 22.8K/Monat4. Aug. 2026
MIT4. Aug. 2026 · Metriken 2.10.0
PyPI · npm
87ExzellentGesundheitsindex
moonshine-ai/moonshine
Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces
C++ · C★ 10.9K↓ 49.9K/Monat28. Aug. 2026
Eigene Lizenz28. Aug. 2026 · Metriken 2.10.0
npm · crates.io
86ExzellentGesundheitsindex
drakulavich/kesha-voice-kit
Give your tools a voice — speech to text and back, 25 languages, up to ~19× faster than Whisper. On your machine.
Rust · TypeScript★ 67↓ 1.029/Monat28. Juli 2026
MIT28. Juli 2026 · Metriken 2.10.0
PyPI
84ExzellentGesundheitsindex
cmusphinx/pocketsphinx
A small speech recognizer
C★ 4.33212. Aug. 2026
Eigene Lizenz12. Aug. 2026 · Metriken 2.10.0
crates.io · Maven · npm +1
84ExzellentGesundheitsindex
k2-fsa/sherpa-onnx
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
C++ · Python★ 14.4K27. Aug. 2026
Apache-2.027. Aug. 2026 · Metriken 2.10.0
PyPI
83ExzellentGesundheitsindex
modelscope/FunClip
FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
Python★ 6.20131. Aug. 2026
MIT31. Aug. 2026 · Metriken 2.10.0
PyPI
81ExzellentGesundheitsindex
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
Python★ 347↓ 244.7K/Monat20. Juli 2026
MIT20. Juli 2026 · Metriken 2.10.0
PyPI
81ExzellentGesundheitsindex
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.4K↓ 1.5M/Monat5. Aug. 2026
BSD-2-Clause5. Aug. 2026 · Metriken 2.10.0
PyPI · npm
78GutGesundheitsindex
bengizmo/voxint
Keine Repository-Beschreibung veröffentlicht.
Python★ 3↓ 2.239/Monat21. Aug. 2026
Apache-2.021. Aug. 2026 · Metriken 2.10.0
npm
78GutGesundheitsindex
speechmatics/speechmatics-js-sdk
Javascript and Typescript SDK for Speechmatics
TypeScript · HTML★ 57↓ 244.1K/Monat26. Juli 2026
MIT26. Juli 2026 · Metriken 2.10.0
crates.io
73GutGesundheitsindex
altunenes/parakeet-rs
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
Rust★ 378↓ 12.5K/Monat29. Juli 2026
MIT29. Juli 2026 · Metriken 2.10.0
PyPI · crates.io
67GutGesundheitsindex
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
C++ · C★ 197↓ 7.559/Monat17. Juli 2026
MIT17. Juli 2026 · Metriken 2.10.0
PyPI
65GutGesundheitsindex
jdepoix/youtube-transcript-api
This is a python API which allows you to get the transcript/subtitles for a given YouTube video. It also works for automatically generated subtitles and it does not require an API key nor a headless browser, like other selenium based solutions do!
Python★ 8.121↓ 28.7M/Monat27. Aug. 2026
MIT27. Aug. 2026 · Metriken 2.10.0
npm · crates.io
62MittelGesundheitsindex
snomiao/otoji
realtime speech ⇄ text, 音を字に — wire mic → STT → translate → speech as an on-device voice graph that spans your devices over WebRTC. Runs in the browser (transformers.js/ONNX/WebGPU); no API keys, nothing leaves the device by default.
TypeScript · Rust★ 2↓ 11.3K/Monat5. Sept. 2026
MIT5. Sept. 2026 · Metriken 2.10.0
Go
59MittelGesundheitsindex
VoiceBlender/voiceblender
A programmable voice platform: SIP and WebRTC call control, multi-party mixing, recording, TTS/STT, and pluggable AI agents (ElevenLabs, VAPI, Pipecat, Deepgram) — all driven through a REST API, webhooks, and a WebSocket event stream
Go★ 9720. Juli 2026
MIT20. Juli 2026 · Metriken 2.10.0
NuGet
57MittelGesundheitsindex
umlx5h/LLPlayer
The media player for language learning, with dual subtitles, AI-generated subtitles, real-time translation, and more!
C#★ 3.9905. Aug. 2026
GPL-3.05. Aug. 2026 · Metriken 2.10.0
Maven
51MittelGesundheitsindex
CrispStrobe/CrisperWeaver
On-device speech-to-text Flutter app powered by CrispASR (ggml / Whisper) — offline, multi-platform, AGPL-3.0.
Dart★ 4130. Juli 2026
AGPL-3.030. Juli 2026 · Metriken 2.10.0
npm
41SchwachGesundheitsindex
surajmandalcell/asrpro
AI powered desktop transcription app with real time speech recognition, file transcription, global hotkeys, and SRT subtitle export.
TypeScript · JavaScript★ 41. Aug. 2026
Keine Lizenz1. Aug. 2026 · Metriken 2.10.0
PyPI · npm · Go +2
28GefährdetGesundheitsindex
alphacep/vosk-api
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
Jupyter Notebook★ 15K↓ 776.1K/Monat12. Aug. 2026
Apache-2.012. Aug. 2026 · Metriken 2.10.0
PyPI
23GefährdetGesundheitsindex
wiseman/py-webrtcvad
Python interface to the WebRTC Voice Activity Detector
C · C++★ 2.492↓ 451.3K/Monat21. Juli 2026
Eigene Lizenz21. Juli 2026 · Metriken 2.10.0
PyPI
11KritischGesundheitsindex
abhirooptalasila/AutoSub
A CLI script to generate subtitle files (SRT/VTT/TXT) for any video using either DeepSpeech or Coqui
Python★ 6514. Aug. 2026
MIT4. Aug. 2026 · Metriken 2.10.0