Stable Diffusion Webui
AUTOMATIC1111
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
33–48 / 113
Results: 113
AUTOMATIC1111
vercel
Enlightened library to convert HTML and CSS to SVG
k2-fsa
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Su…
PaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
supertone-inc
Lightning-Fast, On-Device, Multilingual TTS — running natively via ONNX.
SYSTRAN
Faster Whisper transcription with CTranslate2
abus-aikorea
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTu…
Wan-Video
Wan: Open and Advanced Large-Scale Video Generative Models
speechbrain
openvinotoolkit
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
Zulko
coqui-ai
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
HKUDS
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
duixcom
🚀 Truly open-source AI avatar(digital human) toolkit for offline video generation and digital human cloning.
nari-labs
A TTS model capable of generating ultra-realistic dialogue in one pass.
espnet
End-to-End Speech Processing Toolkit