Baocut
JimLiu
Open-source Agent Skill that drives the BaoCut macOS app CLI (transcribe · subtitle · translate · cut) from Claude Code, Codex, and other agents
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 3 · 16 shown · 411 public entries
Results: 411
JimLiu
Open-source Agent Skill that drives the BaoCut macOS app CLI (transcribe · subtitle · translate · cut) from Claude Code, Codex, and other agents
mltframework
MLT Multimedia Framework
microsoft
A Unified Semi-Supervised Learning Codebase (NeurIPS'22)
OpenShot
OpenShot Video Library (libopenshot) is a free, open-source project dedicated to delivering high quality video editing, animation, and playback solutions to the world. A…
bytedance
SALMONN family: A suite of advanced multi-modal LLMs
FunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
devnen
Self-host the powerful Chatterbox TTS model. This server offers a user-friendly Web UI, flexible API endpoints (incl. OpenAI compatible), predefined voices, voice clonin…
Renumics
Interactively explore unstructured datasets from your dataframe.
Yuan-ManX
Your AI Game Dev Hub. The ultimate resource hub for AI-powered game development tools. Discover cutting-edge LLMs, World Model, Agent, Code, Image, Texture, Shader, 3D M…
myshell-ai
Instant voice cloning by MIT and MyShell. Audio foundation model.
etro-js
Typescript video-editing library for the browser
lhotse-speech
Tools for handling multimodal data in machine learning projects.
diodiogod
A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio…
libAudioFlux
A library for audio and music analysis, feature extraction.
rsxdalv
A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro, OpenVoice, ParlerTTS, Stable Au…
stemrollerapp
Isolate vocals, drums, bass, and other instrumental stems from any song