Todas las etiquetas
Etiqueta del catálogo

#asr

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

14 registros
Con la etiqueta «asr»Ordenado por índice de salud
npm
79Buenoíndice de salud
deepgram/deepgram-js-sdk
Official JavaScript SDK for Deepgram.
TypeScript★ 268↓ 2.4M/mes19 jul 2026
MIT19 jul 2026 · métricas 1.13.0
78Buenoíndice de salud
FluidInference/FluidAudio
Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
Swift★ 246716 jul 2026
Apache-2.016 jul 2026 · métricas 1.13.0
PyPI · npm
78Buenoíndice de salud
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.4K21 jul 2026
MIT21 jul 2026 · métricas 1.13.0
npm · PyPI
77Buenoíndice de salud
alibaba-damo-academy/funasr
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.3K17 jul 2026
MIT17 jul 2026 · métricas 1.13.0
npm · PyPI
77Buenoíndice de salud
alibaba/funasr
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.3K17 jul 2026
MIT17 jul 2026 · métricas 1.13.0
npm
73Buenoíndice de salud
AssemblyAI/assemblyai-node-sdk
The AssemblyAI JavaScript SDK provides an easy-to-use interface for interacting with the AssemblyAI API, which supports async and real-time transcription, audio intelligence models, as well as the latest LeMUR models.
TypeScript★ 76↓ 1.6M/mes15 jul 2026
MIT15 jul 2026 · métricas 1.13.0
PyPI
70Buenoíndice de salud
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.2K21 jul 2026
BSD-2-Clause21 jul 2026 · métricas 1.13.0
crates.io · Maven · npm +1
69Moderadoíndice de salud
k2-fsa/sherpa-onnx
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Android, iOS, HarmonyOS, Raspberry Pi, RISC-V, RK NPU, Axera NPU, Ascend NPU, x86_64 servers, websocket server/client, support 12 programming languages
C++ · Python★ 13.6K17 jul 2026
Apache-2.017 jul 2026 · métricas 1.13.0
PyPI
65Moderadoíndice de salud
jdepoix/youtube-transcript-api
This is a python API which allows you to get the transcript/subtitles for a given YouTube video. It also works for automatically generated subtitles and it does not require an API key nor a headless browser, like other selenium based solutions do!
Python★ 7938↓ 32.2M/mes21 jul 2026
MIT21 jul 2026 · métricas 1.13.0
PyPI · crates.io
61Moderadoíndice de salud
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
C++ · C★ 197↓ 7559/mes17 jul 2026
MIT17 jul 2026 · métricas 1.13.0
npm · crates.io
59Moderadoíndice de salud
snomiao/otoji
realtime speech ⇄ text, 音を字に — wire mic → STT → translate → speech as an on-device voice graph that spans your devices over WebRTC. Runs in the browser (transformers.js/ONNX/WebGPU); no API keys, nothing leaves the device by default.
TypeScript · Rust★ 1↓ 4782/mes15 jul 2026
MIT15 jul 2026 · métricas 1.13.0
Go
55Moderadoíndice de salud
VoiceBlender/voiceblender
A programmable voice platform: SIP and WebRTC call control, multi-party mixing, recording, TTS/STT, and pluggable AI agents (ElevenLabs, VAPI, Pipecat, Deepgram) — all driven through a REST API, webhooks, and a WebSocket event stream
Go★ 9720 jul 2026
MIT20 jul 2026 · métricas 1.13.0
PyPI
33En riesgoíndice de salud
istupakov/onnx-asr
A lightweight Python package for Automatic Speech Recognition using ONNX models
Python★ 347↓ 244.7K/mes20 jul 2026
MIT20 jul 2026 · métricas 1.13.0
PyPI
20Críticoíndice de salud
wiseman/py-webrtcvad
Python interface to the WebRTC Voice Activity Detector
C · C++★ 2492↓ 451.3K/mes21 jul 2026
Licencia propia21 jul 2026 · métricas 1.13.0