Data Efficient Gans
mit-han-lab
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 208
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 208
mit-han-lab
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
elevenlabs
The official Python SDK for the ElevenLabs API.
pydn
A powerful tool that translates ComfyUI workflows into executable Python code.
deepgram
Official Python SDK for Deepgram.
lhotse-speech
Tools for handling multimodal data in machine learning projects.
mayuelala
[AAAI 2024] Follow-Your-Pose: This repo is the official implementation of "Follow-Your-Pose : Pose-Guided Text-to-Video Generation using Pose-Free Videos"
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
MiteshPuthran
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)
mozilla
:robot: :speech_balloon: Deep learning for Text to Speech (Discussion forum: https://discourse.mozilla.org/c/tts)
m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
babysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
AILab-CVC
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
OpenBMB
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
AIDC-AI
🚀 AI 全自动短视频引擎 | AI Fully Automated Short Video Engine
Tencent-Hunyuan
HunyuanVideo: A Systematic Framework For Large Video Generation Model
MoonInTheRiver
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code