π€ Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
$ npx skills add huggingface/diffusersAlternatives
Compare similar skills by workflow fit, trust score, quality, GitHub adoption, maintenance, and install readiness.
Current skill
Kandinsky 2 β multilingual text2image latent diffusion model
π€ Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
$ npx skills add huggingface/diffusersStable Diffusion web UI
$ npx skills add AUTOMATIC1111/stable-diffusion-webuiInvoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.
$ npx skills add invoke-ai/InvokeAIπ AI ε ¨θͺε¨ηθ§ι’εΌζ | AI Fully Automated Short Video Engine
$ npx skills add AIDC-AI/Pixelle-VideoEnlightened library to convert HTML and CSS to SVG
$ npx skills add vercel/satoriVideo, Image and GIF upscale/enlarge(Super-Resolution) and Video frame interpolation. Achieved with Waifu2x, Real-ESRGAN, Real-CUGAN, RTX Video Super Resolution VSR, SRMD, RealSR, Anime4K, RIFE, IFRNet, CAIN, DAIN, and ACNet.
$ npx skills add AaronFeng753/Waifu2x-Extension-GUIFFmpeg libav tutorial - learn how media works from basic to transmuxing, transcoding and more. Translations: πΊπΈ π¨π³ π°π· πͺπΈ π»π³ π§π· π·πΊ
$ npx skills add leandromoreira/ffmpeg-libav-tutorialSANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
$ npx skills add NVlabs/SanaVideo stabilization using gyroscope data
$ npx skills add gyroflow/gyroflowOffline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
$ npx skills add alphacep/vosk-apiUnsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.
$ npx skills add unslothai/unsloth1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
$ npx skills add RVC-Boss/GPT-SoVITSVoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
$ npx skills add OpenBMB/VoxCPMPort of OpenAI's Whisper model in C/C++
$ npx skills add ggml-org/whisper.cppWhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
$ npx skills add m-bain/whisperXTranslate the video from one language to another and embed dubbing & subtitles.
$ npx skills add jianchang512/pyvideotransHow to choose
Use an alternative when it has a clearer install path, higher trust score, fresher maintenance, or better platform fit for your current agent stack. Keep Kandinsky 2 if it already passes your workflow test and repository review.
Next step
Open the compare page, test the install commands in a sandbox, and check each repository before using a skill in production.
Alternatives
Compare similar skills by workflow fit, trust score, quality, GitHub adoption, maintenance, and install readiness.
Current skill
Kandinsky 2 β multilingual text2image latent diffusion model
π€ Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
$ npx skills add huggingface/diffusersStable Diffusion web UI
$ npx skills add AUTOMATIC1111/stable-diffusion-webuiInvoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest AI-driven technologies. The solution offers an industry leading WebUI, and serves as the foundation for multiple commercial products.
$ npx skills add invoke-ai/InvokeAIπ AI ε ¨θͺε¨ηθ§ι’εΌζ | AI Fully Automated Short Video Engine
$ npx skills add AIDC-AI/Pixelle-VideoEnlightened library to convert HTML and CSS to SVG
$ npx skills add vercel/satoriVideo, Image and GIF upscale/enlarge(Super-Resolution) and Video frame interpolation. Achieved with Waifu2x, Real-ESRGAN, Real-CUGAN, RTX Video Super Resolution VSR, SRMD, RealSR, Anime4K, RIFE, IFRNet, CAIN, DAIN, and ACNet.
$ npx skills add AaronFeng753/Waifu2x-Extension-GUIFFmpeg libav tutorial - learn how media works from basic to transmuxing, transcoding and more. Translations: πΊπΈ π¨π³ π°π· πͺπΈ π»π³ π§π· π·πΊ
$ npx skills add leandromoreira/ffmpeg-libav-tutorialSANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
$ npx skills add NVlabs/SanaVideo stabilization using gyroscope data
$ npx skills add gyroflow/gyroflowOffline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
$ npx skills add alphacep/vosk-apiUnsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.
$ npx skills add unslothai/unsloth1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
$ npx skills add RVC-Boss/GPT-SoVITSVoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
$ npx skills add OpenBMB/VoxCPMPort of OpenAI's Whisper model in C/C++
$ npx skills add ggml-org/whisper.cppWhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
$ npx skills add m-bain/whisperXTranslate the video from one language to another and embed dubbing & subtitles.
$ npx skills add jianchang512/pyvideotransHow to choose
Use an alternative when it has a clearer install path, higher trust score, fresher maintenance, or better platform fit for your current agent stack. Keep Kandinsky 2 if it already passes your workflow test and repository review.
Next step
Open the compare page, test the install commands in a sandbox, and check each repository before using a skill in production.