暂未收录效果图
查看技能说明GPT SoVITS
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
OPENAGENTSKILL / DIRECTORY
为下一项任务找到合适的技能。探索适用于 Codex、Claude Code、Cursor 等 Agent 的工具。
760 Skills
搜索结果: 760
暂未收录效果图
查看技能说明RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)

hugohe3
AI generates a real, editable PowerPoint from any document — native shapes & animations, speaker notes voiced as audio narration, and the option to follow your own .pptx…
查看预览 · 2暂未收录效果图
查看技能说明google-ai-edge
暂未收录效果图
查看技能说明huggingface
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and traini…
暂未收录效果图
查看技能说明unslothai
Unsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.
暂未收录效果图
查看技能说明ggml-org
暂未收录效果图
查看技能说明alyssaxuu
暂未收录效果图
查看技能说明alphacep
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
暂未收录效果图
查看技能说明huggingface
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
暂未收录效果图
查看技能说明m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
暂未收录效果图
查看技能说明OpenBMB
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
暂未收录效果图
查看技能说明FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
暂未收录效果图
查看技能说明FunAudioLLM
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
暂未收录效果图
查看技能说明index-tts
暂未收录效果图
查看技能说明screenpipe
YC (S26) | AI that knows what you've seen, said, or heard. Records everything you do, say, hear 24/7, local, private, secure
暂未收录效果图
查看技能说明jianchang512
Translate the video from one language to another and embed dubbing & subtitles.