Cabbage
VideoFlint
A video composition framework build on top of AVFoundation. It's simple to use and easy to extend.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
225โ240 / 478
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 478
VideoFlint
A video composition framework build on top of AVFoundation. It's simple to use and easy to extend.
menyifang
Official implementation of "MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling"
alphacep
Offline speech recognition for Android with Vosk library.
Saik0s
The open-source iOS app that's making quality voice transcription more accessible on mobile devices.
microsoft
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
voice-cloning-app
A Python/Pytorch app for easily synthesising human voices
MiteshPuthran
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)
coqui-ai
๐ A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
lucidrains
Implementation of Video Diffusion Models, Jonathan Ho's new paper extending DDPMs to Video Generation - in Pytorch
ali-vilab
Timestep Embedding Tells: It's Time to Cache for Video Diffusion Model
mayuelala
[AAAI 2024] Follow-Your-Pose: This repo is the official implementation of "Follow-Your-Pose : Pose-Guided Text-to-Video Generation using Pose-Free Videos"
receyuki
A simple standalone viewer for reading prompts from Stable Diffusion generated image outside the webui.
CSTR-Edinburgh
This is now the official location of the Merlin project.
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.