Data Efficient Gans
mit-han-lab
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1โ16 / 20
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 20
mit-han-lab
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
jy0205
[ICLR 2025] Pyramidal Flow Matching for Efficient Video Generative Modeling
mozilla
:robot: :speech_balloon: Deep learning for Text to Speech (Discussion forum: https://discourse.mozilla.org/c/tts)
X-GenGroup
A unified framework for easy reinforcement learning in Flow-Matching models
FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
zzw922cn
End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow
Breakthrough
:movie_camera: Python and OpenCV-based scene cut/transition detection program & library.
shivammehta25
[ICASSP 2024] ๐ต Matcha-TTS: A fast TTS architecture with conditional flow matching
lhotse-speech
Tools for handling multimodal data in machine learning projects.
MiteshPuthran
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Aratako
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
Blaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
AILab-CVC
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
AceDataCloud
Consumer AI app for chat, image generation, video generation, and music creation powered by Ace Data Cloud APIs.