Data Efficient Gans
mit-han-lab
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 19
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 19
mit-han-lab
[NeurIPS 2020] Differentiable Augmentation for Data-Efficient GAN Training
kaltura
The Kaltura Platform Backend. To install Kaltura, visit the install packages repository.
lhotse-speech
Tools for handling multimodal data in machine learning projects.
zzw922cn
End-to-end Automatic Speech Recognition for Madarian and English in Tensorflow
MiteshPuthran
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)
sandrohanea
Whisper.net. Speech to text made simple using Whisper Models
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
toki-plus
全自动短视频搬运工具,支持自动下载、去重、AI生成标题+标签、上传,可二开扩展至多平台,例如:TikTok->视频号/抖音/小红书、抖音->TikTok/视频号/小红书......video-processing, automation, tiktok, selenium, pyqt5, ffmpeg, bot, data-scrapi…
AceDataCloud
Consumer AI app for chat, image generation, video generation, and music creation powered by Ace Data Cloud APIs.
opencast
The free and open source solution for automated video capture and distribution at scale.
abhiTronix
A High-performance cross-platform Video Processing Python framework powerpacked with unique trailblazing features :fire:
AILab-CVC
VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models
yuanzhongqiao
短剧平台 AI Short Film Motion Comic Generation Platform Industrial AI Motion Comic & Video Workbench
OpenGVLab
InternGPT (iGPT) is an open source demo platform where you can easily showcase your AI models. Now it supports DragGAN, ChatGPT, ImageBind, multimodal chat like GPT-4, S…
yeyupiaoling
Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inf…