elevenlabs-tts
MengTo
Generate ElevenLabs text-to-speech audio from scripts or inline text using local voice profiles. Use when the user asks for ElevenLabs, text-to-speech, TTS, narration, v…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 17 · 16 shown · 1,112 public entries
Results: 1112
MengTo
Generate ElevenLabs text-to-speech audio from scripts or inline text using local voice profiles. Use when the user asks for ElevenLabs, text-to-speech, TTS, narration, v…
SamurAIGPT
Remove Seedance 2.0 watermark (AI生成) from videos automatically. No GPU required. Free open-source tool.
VideoFlint
A video composition framework build on top of AVFoundation. It's simple to use and easy to extend.
menyifang
Official implementation of "MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling"
Devin-AXIS
Reuse approved HyperFrames blocks and components inside the active iPollo Video project.
crocidb
bulletty is a pretty feed reader for the terminal that stores the articles as Markdown
kaltura
The Kaltura Platform Backend. To install Kaltura, visit the install packages repository.
mbzuai-oryx
[ACL 2024 🔥] Video-ChatGPT is a video conversation model capable of generating meaningful conversation about videos. It combines the capabilities of LLMs with a pretrai…
Anil-matcha
AI short drama & micro-drama video generator — turns any idea into a complete short-form drama using multi-agent AI pipeline (screenwriter → storyboard → frames → video)…
infiniV
Local voice dictation and meeting recorder for Windows + Linux. Hold a hotkey to dictate, or record long-form meetings with system audio. Whisper transcription, bring-yo…
jaxxchen003
Portable Codex skill for auditable, rights-aware Chinese book-review short-video workflows.
wenhaochai
[ICCV 2023] StableVideo: Text-driven Consistency-aware Diffusion Video Editing
ORB-HD
Video anonymization by face detection
amanchadha
iSeeBetter: Spatio-Temporal Video Super Resolution using Recurrent-Generative Back-Projection Networks | Python3 | PyTorch | GANs | CNNs | ResNets | RNNs | Published in…
MiteshPuthran
The neural network model is capable of detecting five different male/female emotions from audio speeches. (Deep Learning, NLP, Python)
wanshuiyin
Agentic, long-horizon visual generation: a fuzzy story → a cross-model-audited image-based movie. Brings ARIS's research-wiki + multi-agent debate to multimodal generati…