No visual example yet
Explore the skillFunClip
modelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
OPENAGENTSKILL / DIRECTORY
다음 작업에 맞는 스킬을 찾아보세요. Codex, Claude Code, Cursor 등을 지원합니다.
83 Skills
검색 결과: 83
No visual example yet
Explore the skillmodelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
No visual example yet
Explore the skilldenizsafak
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
No visual example yet
Explore the skillgoogle-ai-edge
Cross-platform, customizable ML solutions for live and streaming media.
No visual example yet
Explore the skillRVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
No visual example yet
Explore the skillm-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
No visual example yet
Explore the skillbabysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
No visual example yet
Explore the skillggml-org
Port of OpenAI's Whisper model in C/C++
No visual example yet
Explore the skillAIDC-AI
🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine
No visual example yet
Explore the skillTencent-Hunyuan
HunyuanVideo: A Systematic Framework For Large Video Generation Model
No visual example yet
Explore the skillKlingAIResearch
No visual example yet
Explore the skilljianchang512
Translate the video from one language to another and embed dubbing & subtitles.
No visual example yet
Explore the skillWan-Video
Wan: Open and Advanced Large-Scale Video Generative Models
No visual example yet
Explore the skillopenvinotoolkit
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
No visual example yet
Explore the skillZulko
No visual example yet
Explore the skillHKUDS
"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"
No visual example yet
Explore the skillnari-labs
A TTS model capable of generating ultra-realistic dialogue in one pass.