Vosk API
alphacep
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 19
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 19
alphacep
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
bloc97
A High-Quality Real Time Upscaler for Anime Video
mravanelli
pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch, while feature extraction, label…
MiteshPuthran
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)
invoke-ai
Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest…
pannous
🎙Speech recognition using the tensorflow deep learning framework, sequence-to-sequence neural networks
elevenlabs
The official Python SDK for the ElevenLabs API.
gradio-app
The python library for real-time communication
kan-bayashi
Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch
lucidrains
Implementation of Video Diffusion Models, Jonathan Ho's new paper extending DDPMs to Video Generation - in Pytorch
syhw
Attempt at tracking states of the arts and recent results (bibliography) on speech recognition.
Azure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
amanchadha
iSeeBetter: Spatio-Temporal Video Super Resolution using Recurrent-Generative Back-Projection Networks | Python3 | PyTorch | GANs | CNNs | ResNets | RNNs | Published in…
taco-group
SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation