video-voiceover
zenstory-ai
把带时间戳的 narration.json 合成为中文解说音频。使用 MiMo TTS(mimo-v2.5-tts)或 Fish Audio(s2.1-pro-free)逐段生成语音, 按时间窗动态适配语速并处理响度;输入输出时间线上的旁白,产出 tts_segments 与 tts_meta.json。 旧版直接剪辑路径也可显式传入…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 31 · 16 shown · 1,117 public entries
Results: 1117
zenstory-ai
把带时间戳的 narration.json 合成为中文解说音频。使用 MiMo TTS(mimo-v2.5-tts)或 Fish Audio(s2.1-pro-free)逐段生成语音, 按时间窗动态适配语速并处理响度;输入输出时间线上的旁白,产出 tts_segments 与 tts_meta.json。 旧版直接剪辑路径也可显式传入…
wildminder
ComfyUI custom node for the VibeVoice TTS. Expressive, long-form, multi-speaker conversational audio
JunyaoHu
You can easily calculate FVD, PSNR, SSIM, LPIPS for evaluating the quality of generated or predicted videos.
GuanYixuan
A lightweight, flexible, and easy-to-use Python tool for generating and exporting CapCut drafts to build fully automated video editing/remix pipelines! Another similar p…
VideoVerses
Let's finetune video generation models!
yanhua1010
把已确认的母题或文案转成可直接拍摄、录屏或交给视频工具制作的短视频方案,也支持在用户确认肖像与声音权利并完成平台手动上传后制作数字人视频。用于视频号、抖音、小红书视频和其他竖屏短视频的口播稿、前 3 秒钩子、分镜、字幕、录屏清单、封面、话题、BGM 建议、数字人制片包和发布文案。
Affitor
Write short-form video scripts for TikTok, Instagram Reels, and YouTube Shorts that promote affiliate products with strong hooks, demos, and CTAs. Use this skill when th…
ID-LoRA
Custom ComfyUI node for generating videos with audio-visual identity based on a reference voice and image
bcmi
[ICCV 2025] Light-A-Video: Training-free Video Relighting via Progressive Light Fusion
Fantasy-AMAP
FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers
SonarSonic
DrawingBotV3 is a software for converting images into vector art
SWHL
🎦 Extract video hard subtitles and automatically generate corresponding srt files.
caiyuanhao1998
OmniVCus: Feedforward Subject-driven Video Customization with Multimodal Control Conditions (NeurIPS 2025)
baaivision
[ICLR 2026] 🐻 Uniform Discrete Diffusion with Metric Path for Video Generation
double22a
The dataset of Speech Recognition
asjqkkkk
📖Rendering markdown by flutter!Welcome for pr and issue.