Todas las etiquetas
Etiqueta del catálogo

#speech

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

39 registros
Con la etiqueta «speech»Ordenado por índice de salud
PyPI
98Excepcionalíndice de salud
huggingface/datasets
🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools
Python★ 21.8K↓ 146M/mes5 ago 2026
Apache-2.05 ago 2026 · métricas 2.10.0
PyPI
97Excepcionalíndice de salud
modelscope/modelscope
ModelScope: bring the notion of Model-as-a-Service to life.
Python★ 9112↓ 5.5M/mes28 ago 2026
Apache-2.028 ago 2026 · métricas 2.10.0
Packagist
96Excepcionalíndice de salud
googleapis/google-cloud-php
Google Cloud Client Library for PHP
PHP★ 1181↓ 189.3K/mes22 ago 2026
Apache-2.022 ago 2026 · métricas 2.10.0
PyPI · npm
92Excelenteíndice de salud
Picovoice/porcupine
On-device wake word detection powered by deep learning
Python · TypeScript · Swift★ 4911↓ 773K/mes12 ago 2026
Apache-2.012 ago 2026 · métricas 2.10.0
PyPI
90Excelenteíndice de salud
pytorch/audio
Data manipulation and transformation for audio signal processing, powered by PyTorch
Python★ 290617 jul 2026
BSD-2-Clause17 jul 2026 · métricas 2.10.0
PyPI
87Excelenteíndice de salud
mbailey/voicemode
Natural voice conversations with Claude Code
Python★ 1308↓ 22.8K/mes4 ago 2026
MIT4 ago 2026 · métricas 2.10.0
npm
87Excelenteíndice de salud
microsoft/cognitive-services-speech-sdk-js
Microsoft Azure Cognitive Services Speech SDK for JavaScript
TypeScript★ 322↓ 1.2M/mes21 jul 2026
Licencia propia21 jul 2026 · métricas 2.10.0
PyPI · npm
87Excelenteíndice de salud
moonshine-ai/moonshine
Very low latency speech to text, intent recognition, and text to speech, for building voice agents and interfaces
C++ · C★ 10.9K↓ 49.9K/mes28 ago 2026
Licencia propia28 ago 2026 · métricas 2.10.0
PyPI
86Excelenteíndice de salud
Uberi/speech_recognition
Speech recognition module for Python, supporting several engines and APIs, online and offline.
Python★ 8987↓ 9.2M/mes27 ago 2026
BSD-3-Clause27 ago 2026 · métricas 2.10.0
PyPI
84Excelenteíndice de salud
cmusphinx/pocketsphinx
A small speech recognizer
C★ 433212 ago 2026
Licencia propia12 ago 2026 · métricas 2.10.0
PyPI
84Excelenteíndice de salud
lhotse-speech/lhotse
Tools for handling multimodal data in machine learning projects.
Python★ 1146↓ 1.4M/mes13 ago 2026
Apache-2.013 ago 2026 · métricas 2.10.0
npm · PyPI · crates.io
83Excelenteíndice de salud
decibri/decibri
Cross-platform audio capture, playback, and voice activity detection for Python, Rust, and Node.js powered by a single Rust core.
Rust · Python · JavaScript★ 23↓ 19.3K/mes23 jul 2026
Apache-2.023 jul 2026 · métricas 2.10.0
npm · crates.io
83Excelenteíndice de salud
silverstein/minutes
Every meeting, every idea, every voice note — searchable by your AI. Open-source, privacy-first conversation memory layer.
Rust★ 1372↓ 5342/mes23 jul 2026
MIT23 jul 2026 · métricas 2.10.0
PyPI
81Excelenteíndice de salud
m-bain/whisperX
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
Python★ 23.4K↓ 1.5M/mes5 ago 2026
BSD-2-Clause5 ago 2026 · métricas 2.10.0
NuGet
80Excelenteíndice de salud
shinyorg/shiny
.NET Framework for Backgrounding & Device Hardware Services (iOS, MacCatalyst, Android, & Windows)
C#★ 158123 ago 2026
MIT23 ago 2026 · métricas 2.10.0
Packagist
80Excelenteíndice de salud
symfony/ai-platform
PHP library for interacting with AI platform provider.
PHP★ 54↓ 212.3K/mes6 sept 2026
MIT6 sept 2026 · métricas 2.10.0
PyPI
78Buenoíndice de salud
CUNY-CL/wikipron
Massively multilingual pronunciation mining
Python★ 372↓ 3440/mes20 ago 2026
Apache-2.020 ago 2026 · métricas 2.10.0
PyPI · npm
78Buenoíndice de salud
Picovoice/orca
On-device streaming text-to-speech engine powered by deep learning
Python · C# · TypeScript★ 144↓ 11K/mes18 ago 2026
Apache-2.018 ago 2026 · métricas 2.10.0
PyPI
78Buenoíndice de salud
snakers4/silero-vad
Silero VAD: pre-trained enterprise-grade Voice Activity Detector
Python · Jupyter Notebook★ 992712 ago 2026
MIT12 ago 2026 · métricas 2.10.0
npm
78Buenoíndice de salud
speechmatics/speechmatics-js-sdk
Javascript and Typescript SDK for Speechmatics
TypeScript · HTML★ 57↓ 244.1K/mes26 jul 2026
MIT26 jul 2026 · métricas 2.10.0
PyPI
77Buenoíndice de salud
YannickJadoul/Parselmouth
Praat in Python, the Pythonic way
C++ · Python★ 1278↓ 380.6K/mes13 ago 2026
GPL-3.013 ago 2026 · métricas 2.10.0
crates.io
77Buenoíndice de salud
ai-coustics/aic-sdk-rs
Rust Wrapper for the ai-coustics SDK
Rust · C★ 4↓ 13.2K/mes5 sept 2026
Apache-2.05 sept 2026 · métricas 2.10.0
PyPI
75Buenoíndice de salud
nateshmbhat/pyttsx3
Offline Text To Speech synthesis for python
Python★ 2530↓ 970.4K/mes13 ago 2026
MPL-2.013 ago 2026 · métricas 2.10.0
crates.io
73Buenoíndice de salud
altunenes/parakeet-rs
very fast speech-to-text, diarization, streaming (even in CPU) with NVIDIA Parakeet in Rust
Rust★ 378↓ 12.5K/mes29 jul 2026
MIT29 jul 2026 · métricas 2.10.0
PyPI · crates.io
67Buenoíndice de salud
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
C++ · C★ 197↓ 7559/mes17 jul 2026
MIT17 jul 2026 · métricas 2.10.0
npm
65Buenoíndice de salud
hexinmiao96/vue-ai-hooks
El repositorio no publica descripción.
TypeScript · JavaScript★ 1↓ 3121/mes31 jul 2026
MIT31 jul 2026 · métricas 2.10.0
npm
65Buenoíndice de salud
kuralle/syrinx
Open-source voice agent engine in TypeScript — provider-neutral STT/LLM/TTS pipeline, telephony, barge-in; runs on Node & Cloudflare Workers
TypeScript★ 0↓ 11.3K/mes26 ago 2026
MIT26 ago 2026 · métricas 2.10.0
PyPI
63Moderadoíndice de salud
SYSTRAN/faster-whisper
Faster Whisper transcription with CTranslate2
Python★ 24.7K5 ago 2026
MIT5 ago 2026 · métricas 2.10.0
npm · crates.io
63Moderadoíndice de salud
ai-coustics/aic-sdk-node
Node.js bindings for ai-coustics speech enhancement SDK
JavaScript · Rust★ 5↓ 5276/mes31 jul 2026
Apache-2.031 jul 2026 · métricas 2.10.0
npm · crates.io
62Moderadoíndice de salud
snomiao/otoji
realtime speech ⇄ text, 音を字に — wire mic → STT → translate → speech as an on-device voice graph that spans your devices over WebRTC. Runs in the browser (transformers.js/ONNX/WebGPU); no API keys, nothing leaves the device by default.
TypeScript · Rust★ 2↓ 11.3K/mes5 sept 2026
MIT5 sept 2026 · métricas 2.10.0