LightX2V
ModelTC
Lightweight Image Video Action Generation Inference Framework
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
17–32 / 85
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 85
ModelTC
Lightweight Image Video Action Generation Inference Framework
EvolvingLMMs-Lab
A simple, unified multimodal models training engine. Lean, flexible, and built for hacking at scale.
lone-cloud
A desktop app for running Large Language Models locally.
yl4579
StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
Tencent-Hunyuan
HunyuanVideo: A Systematic Framework For Large Video Generation Model
junyanz
Image-to-Image Translation in PyTorch
junyanz
Software that can generate photos from paintings, turn horses into zebras, perform style transfer, and more.
vllm-project
A framework for efficient model inference with omni-modality models
phillipi
Image-to-image translation with conditional adversarial nets
fudan-generative-vision
[ECCV 2024] Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
PaddlePaddle
PaddlePaddle GAN library, including lots of interesting applications like First-Order motion transfer, Wav2Lip, picture repair, image editing, photo2cartoon, image style…
thu-ml
TurboDiffusion: 100–200× Acceleration for Video Diffusion Models
PKU-YuanGroup
Helios: Real Real-Time Long Video Generation Model
pnnbao97
Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng…
MoonInTheRiver
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code