Data Efficient Gans
mit-han-lab
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 16
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 16
mit-han-lab
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
mozilla
:robot: :speech_balloon: Deep learning for Text to Speech (Discussion forum: https://discourse.mozilla.org/c/tts)
cboard-org
Augmentative and Alternative Communication (AAC) system with text-to-speech for the browser
FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
Breakthrough
:movie_camera: Python and OpenCV-based scene cut/transition detection program & library.
zzw922cn
End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow
lhotse-speech
Tools for handling multimodal data in machine learning projects.
MiteshPuthran
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Blaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
AILab-CVC
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
AceDataCloud
Consumer AI app for chat, image generation, video generation, and music creation powered by Ace Data Cloud APIs.
toki-plus
全自动短视频搬运工具,支持自动下载、去重、AI生成标题+标签、上传,可二开扩展至多平台,例如:TikTok->视频号/抖音/小红书、抖音->TikTok/视频号/小红书......video-processing, automation, tiktok, selenium, pyqt5, ffmpeg, bot, data-scrapi…
yeyupiaoling
Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inf…
Eyeline-Labs
Official code, models, and data for Vista4D: Video Reshooting with 4D Point Clouds (CVPR 2026 Highlight)