StableAvatar
Francis-Rings
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-proβ¦
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
81β96 / 398
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 398
Francis-Rings
We present StableAvatar, the first end-to-end video diffusion transformer, which synthesizes infinite-length high-quality audio-driven avatar videos without any post-proβ¦
ekwek1
Soprano: Instant, Ultra-Realistic Text-to-Speech
harskish
Discovering Interpretable GAN Controls [NeurIPS 2020]
ali-vilab
Official implementations for paper: DreamTalk: When Expressive Talking Head Generation Meets Diffusion Probabilistic Models
varunshenoy
An extensible, easy-to-use, and portable diffusion web UI π¨βπ¨
nari-labs
TTS model capable of streaming conversational audio in realtime.
thu-ml
Official code of Motus: A Unified Latent Action World Model
VideoFlint
A video composition framework build on top of AVFoundation. It's simple to use and easy to extend.
alphacep
Offline speech recognition for Android with Vosk library.
voice-cloning-app
A Python/Pytorch app for easily synthesising human voices
lucidrains
Implementation of Video Diffusion Models, Jonathan Ho's new paper extending DDPMs to Video Generation - in Pytorch
mayuelala
[AAAI 2024] Follow-Your-Pose: This repo is the official implementation of "Follow-Your-Pose : Pose-Guided Text-to-Video Generation using Pose-Free Videos"
snap-research
Code for Motion Representations for Articulated Animation paper
huanngzh
[ICCV 2025] Official impl. of "MV-Adapter: Multi-view Consistent Image Generation Made Easy"
sdkcarlos
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your own siri,google now or cortana with Google Chrome within yourβ¦