No visual example yet
Explore the skillGyroflow
gyroflow
Video stabilization using gyroscope data
OPENAGENTSKILL / DIRECTORY
Finde den passenden Skill für deine nächste Aufgabe mit Codex, Claude Code, Cursor und mehr.
113 Skills
Ergebnisse: 113
No visual example yet
Explore the skillgyroflow
Video stabilization using gyroscope data
No visual example yet
Explore the skillUberi
Speech recognition module for Python, supporting several engines and APIs, online and offline.
No visual example yet
Explore the skillrany2
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key
No visual example yet
Explore the skillcamenduru
stable diffusion webui colab
No visual example yet
Explore the skillkaldi-asr
kaldi-asr/kaldi is the official location of the Kaldi project.
No visual example yet
Explore the skillVectorSpaceLab
OmniGen: Unified Image Generation. https://arxiv.org/pdf/2409.11340
No visual example yet
Explore the skillNVlabs
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
No visual example yet
Explore the skilldebpalash
The open-source ElevenLabs alternative for local voice cloning, design, create, dubbing and dictation Desktop App
No visual example yet
Explore the skillBlaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
No visual example yet
Explore the skillzai-org
text and image to video generation: CogVideoX (2024) and CogVideo (ICLR 2023)
No visual example yet
Explore the skillopen-mmlab
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get…
No visual example yet
Explore the skillTalAter
💬 Speech recognition for your site
No visual example yet
Explore the skillargmaxinc
On-device Speech AI for Apple Silicon
No visual example yet
Explore the skillleejet
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++
No visual example yet
Explore the skillsnakers4
Silero Models: pre-trained text-to-speech models made embarrassingly simple
No visual example yet
Explore the skillnl8590687
A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统