OmniShow
Correr-Zhou
[ICML 2026] ByteDance's All-in-One Video Generation Model for Human-Object Interaction Video Generation
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
257–272 / 477
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 477
Live data is unavailable. A saved snapshot may be shown; check the source before use.
Correr-Zhou
[ICML 2026] ByteDance's All-in-One Video Generation Model for Human-Object Interaction Video Generation
PunithVT
🎭 AI Avatar / digital human platform — upload a photo, clone a voice, talk to any face in real time with lip-sync video. Open-source, self-hosted. Claude · Whisper · Ch…
large-performance-model
LPM 1.0: Video-based Character Performance Model
devnen
Self-host the powerful Dia TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), support for SafeTensors/BF16, voice cl…
google-marketing-solutions
Recrafting Video Ads with Generative AI
Saganaki22
OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and multi-speaker dialogue
ChrisChen667788
Multi-agent AI pipeline that turns one line of text into a finished short-form drama: script, cinematic storyboards, character-consistent video. Provider-agnostic (OpenA…
soupslurpr
Private and on-device speech recognition keyboard and service for Android.
A-cat-with-carrots
OnlyShot · 神仙鱼 AI 短剧工业流水线 Claude Skill — 一句话 → 剧本/分镜/分镜图/视频/剪辑 五步 → 可发红果抖音的完整短剧。集成即梦/Seedance,17 失败模式沉淀。
devnen
Self-host the ultra-lightweight Kitten TTS model with this enhanced API server with an intuitive Web UI, large text processing for audiobooks, and GPU acceleration.
kapi2800
Run Qwen3-TTS text-to-speech locally on Mac (M1/M2/M3/M4). Voice cloning, voice design, custom voices. 100% offline using MLX.
jeffstric
ZhiJuTong (ZJT) is an AI-powered, open-source platform specifically designed for creating professional short dramas. It automates the entire production pipeline, from sc…
CIntellifusion
Official Implementation of MultiWorld: Scalable Multi-Agent Multi-View Video World Models
OSideMedia
Claude AI skill for cinematic Higgsfield AI prompts — 20 sub-skills covering Cinema Studio 2.5/3.0/3.5, MCSLA formula, Soul ID character consistency, Seedance 2.0 prompt…
aHapBean
[NeurIPS 2025] VideoREPA: Learning Physics for Video Generation through Relational Alignment with Foundation Models