DragGAN
OpenGVLab
Unofficial Implementation of DragGAN - "Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold" (DragGAN 全功能实现,在线Demo,本地部署试用,代码、模型已全部开源,支持W…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
113–128 / 429
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 429
OpenGVLab
Unofficial Implementation of DragGAN - "Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold" (DragGAN 全功能实现,在线Demo,本地部署试用,代码、模型已全部开源,支持W…
Stability-AI
StableSwarmUI, A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and extensibility.
thu-ml
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, i…
sanchit-gandhi
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
pydn
A powerful tool that translates ComfyUI workflows into executable Python code.
SkyworkAI
Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory
Tencent-Hunyuan
HunyuanImage-3.0: A Powerful Native Multimodal Model for Image Generation
OpenShot
OpenShot Video Library (libopenshot) is a free, open-source project dedicated to delivering high quality video editing, animation, and playback solutions to the world. A…
mkiol
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
Purfview
Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
fudan-generative-vision
[ECCV 2024] Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance
Picsart-AI-Research
[ICCV 2023 Oral] Text-to-Image Diffusion Models are Zero-Shot Video Generators
ali-vilab
Official implementations for paper: Anydoor: zero-shot object-level image customization
bytedance
SALMONN family: A suite of advanced multi-modal LLMs
metavoiceio
Foundational model for human-like, expressive TTS
huggingface
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.