No visual example yet
Explore the skillModelscope
modelscope
ModelScope: bring the notion of Model-as-a-Service to life.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–16 / 25
Results: 25
No visual example yet
Explore the skillmodelscope
ModelScope: bring the notion of Model-as-a-Service to life.
No visual example yet
Explore the skilldexhunter
Write precise Seedance 2.0 prompts for multimodal video, camera movement, editing, music, and product storytelling.
No visual example yet
Explore the skillTencent-Hunyuan
HunyuanVideo: A Systematic Framework For Large Video Generation Model
No visual example yet
Explore the skillvllm-project
A framework for efficient model inference with omni-modality models
No visual example yet
Explore the skillPKU-YuanGroup
Helios: Real Real-Time Long Video Generation Model
No visual example yet
Explore the skillFunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
No visual example yet
Explore the skillbytedance
A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.
No visual example yet
Explore the skillRimagination
A hardware-aware Codex skill for local MiniMax H3 video generation through ComfyUI
No visual example yet
Explore the skillRVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
No visual example yet
Explore the skillggml-org
Port of OpenAI's Whisper model in C/C++
No visual example yet
Explore the skillFunAudioLLM
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
No visual example yet
Explore the skillPaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
No visual example yet
Explore the skillnari-labs
A TTS model capable of generating ultra-realistic dialogue in one pass.
No visual example yet
Explore the skillmyshell-ai
Instant voice cloning by MIT and MyShell. Audio foundation model.
No visual example yet
Explore the skillGetStream
Open Vision Agents by Stream. Build voice and vision agents quickly with any model or video provider. Uses Stream's edge network for ultra-low latency.
No visual example yet
Explore the skillleejet
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++