DreamLayer
TheDesignFounder
Benchmark diffusion models faster. Automate evals, seeds, and metrics for reproducible results.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
129–144 / 289
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 289
TheDesignFounder
Benchmark diffusion models faster. Automate evals, seeds, and metrics for reproducible results.
Shilin-LU
[ICLR 2025] "Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances" (Official Implementation)
ChaitanyaEswarRajeshJakki
A fully autonomous AI Agent/Python pipeline that utilizes Large Language Models (LLMs) like Gemini to generate content, produce videos, and automatically upload educatio…
mozilla
:robot: :speech_balloon: Deep learning for Text to Speech (Discussion forum: https://discourse.mozilla.org/c/tts)
neggles
a CLI utility/library for AnimateDiff stable diffusion generation
netease-youdao
Confucius4-TTS: a Multilingual and Cross-Lingual Zero-Shot TTS Engine
PaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
abus-aikorea
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTu…
AIDC-AI
Ovis-Image is a 7B text-to-image model specifically optimized for high-quality text rendering, designed to operate efficiently under stringent computational constraints.
herimor
VoXtream is a Full-Stream Zero-shot TTS model with Extremely Low Latency and Speaking rate Control
junyanz
Image-to-Image Translation in PyTorch