OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
257–272 / 478
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 478
omerbt
Official Pytorch Implementation for "MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation" presenting "MultiDiffusion" (ICML 2023)
showlab
[ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.
videoflow
Python framework that facilitates the quick development of complex video analysis applications and other series-processing based applications in a multiprocessing enviro…
TensorSpeech
:zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords
Aratako
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
alesaccoia
Near-Realtime audio transcription using self-hosted Whisper and WebSocket in Python/JS
k2-fsa
Speech-to-text server framework with next-gen Kaldi
BinWang28
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
Vonage
Vonage REST API client for PHP. API support for SMS, Voice, Text-to-Speech, Numbers, Verify (2FA) and more.
sandrohanea
Whisper.net. Speech to text made simple using Whisper Models
soniqo
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
TheStageAI
Optimized Whisper models for streaming and on-device use
EDDiscovery
Captains log and 3d star map for Elite Dangerous
AutoArk
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!