Cosmos Drive Dreams
nv-tlabs
Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
129–144 / 446
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 446
nv-tlabs
Cosmos-Drive-Dreams: Scalable Synthetic Driving Data Generation with World Foundation Models
SOTAMak1r
A Unified Visual Generator with Interleaved OmniModal Context
unslothai
Unsloth Studio is a web UI for training and running open models like Gemma 4, Qwen3.6, DeepSeek, gpt-oss locally.
FunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
wernerturing
Detect and remove logos from videos, even if they change position several times
huggingface
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
OpenBMB
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
invoke-ai
Invoke is a leading creative engine for Stable Diffusion models, empowering professionals, artists, and enthusiasts to generate and create visual media using the latest…
dectalk
Modern builds for the 90s/00s DECtalk text-to-speech application.
Migushthe2nd
A simple Azure Speech Service module that uses the Microsoft Edge Read Aloud API. https://www.npmjs.com/package/msedge-tts
AaronFeng753
Video, Image and GIF upscale/enlarge(Super-Resolution) and Video frame interpolation. Achieved with Waifu2x, Real-ESRGAN, Real-CUGAN, RTX Video Super Resolution VSR, SRM…
index-tts
An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
TheDesignFounder
Benchmark diffusion models faster. Automate evals, seeds, and metrics for reproducible results.
alphacep
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
manycore-research
[3DV 2026] SpatialGen: Layout-guided 3D Indoor Scene Generation