Rhino
Picovoice
On-device Speech-to-Intent engine powered by deep learning
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Browse all public registry entries, one page at a time. Counts include resources; MCP-only resources are omitted from displayed skills. Listing is not a safety or compatibility guarantee.
Page 16 · 16 shown · 828 public entries
Results: 828
Picovoice
On-device Speech-to-Intent engine powered by deep learning
fluxions-ai
Real-time voice assistant — WebRTC streaming, faster-whisper ASR, local LLM, Vui Nano (300M) TTS. OpenAI Realtime API compatible. Voice cloning, barge-in, ~9× realtime o…
NVlabs
rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale
coqui-ai
🐸STT - The deep learning toolkit for Speech-to-Text. Training and deploying STT models has never been so easy.
marytts
MARY TTS -- an open-source, multilingual text-to-speech synthesis system written in pure java
PurpleDoubleD
Local AI desktop app — chat, agent mode, image gen, video gen. Supports Ollama, Gemma 4, Llama, Qwen, OpenAI, Anthropic. Single .exe, no Docker.
nvidia-cosmos
Cosmos-Transfer2.5, built on top of Cosmos-Predict2.5, produces high-quality world simulations conditioned on multiple spatial control inputs.
Finrandojin
AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B,…
taco-group
SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation
Blaizzy
A modular Swift SDK for audio processing with MLX on Apple Silicon
rapidaai
Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channe…
Picovoice
On-device streaming speech-to-text engine powered by deep learning
ErickWendel
JS Expert Week 8.0 - 🎥Pre processing videos before uploading in the browser 😏
pnlpal
📚 A customizable dictionary extension that supports double-click lookups in 20+ languages, 1000+ dictionaries, text-to-speech, translation and Anki integration.
mravanelli
pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch, while feature extraction, label…
jamsch
Speech Recognition for React Native Expo projects
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
RedSkill · 流白Livo · 1.0.0
Describe a rain curtain, growing flowering branches or a flock of swallows. This RedSkill package adapts three p5.js templates into sketch.js code with custom colors, density and speed.
Noncommercial use only · Local runtime recording · Use on the source platform