HunyuanCustom
Tencent-Hunyuan
HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
225β240 / 478
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 478
Tencent-Hunyuan
HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation
wladradchenko
Wunjo CE: Face Swap, Lip Sync, Control Remove Objects & Text & Background, Restyling, Audio Separator, Clone Voice, Video Generation. Open Source, Local & Free.
varunshenoy
An extensible, easy-to-use, and portable diffusion web UI π¨βπ¨
kan-bayashi
Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch
nari-labs
TTS model capable of streaming conversational audio in realtime.
thu-ml
Official code of Motus: A Unified Latent Action World Model
sooftware
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
VideoFlint
A video composition framework build on top of AVFoundation. It's simple to use and easy to extend.
menyifang
Official implementation of "MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling"
alphacep
Offline speech recognition for Android with Vosk library.
Saik0s
The open-source iOS app that's making quality voice transcription more accessible on mobile devices.
microsoft
Unified-Modal Speech-Text Pre-Training for Spoken Language Processing
voice-cloning-app
A Python/Pytorch app for easily synthesising human voices
MiteshPuthran
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)