GPT SoVITS
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
17–32 / 108
Results: 108
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
catboost
A fast, scalable, high performance Gradient Boosting on Decision Trees library, used for ranking, classification, regression and other machine learning tasks for Python,…
py-why
DoWhy is a Python library for causal inference that supports explicit modeling and testing of causal assumptions. DoWhy is based on a unified language for causal inferen…
m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
stanfordnlp
Stanford NLP Python library for tokenization, sentence segmentation, NER, and parsing of many human languages
Blaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
babysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
unslothai
Unsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.
ggml-org
Port of OpenAI's Whisper model in C/C++
FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
huggingface
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
OpenBMB
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
invoke-ai
Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest…
AIDC-AI
🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine
Tencent-Hunyuan
HunyuanVideo: A Systematic Framework For Large Video Generation Model
AaronFeng753
Video, Image and GIF upscale/enlarge(Super-Resolution) and Video frame interpolation. Achieved with Waifu2x, Real-ESRGAN, Real-CUGAN, RTX Video Super Resolution VSR, SRM…