Vox AI Motion Graphics Generator
Anil-matcha
🎬 Turn any topic into a finished Vox-style paper-collage explainer / motion graphics video — script, collage keyframes, animation, voice-over, music & captions, all aut…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
273–288 / 477
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 477
Anil-matcha
🎬 Turn any topic into a finished Vox-style paper-collage explainer / motion graphics video — script, collage keyframes, animation, voice-over, music & captions, all aut…
World-In-World
Code implementation of the paper "World-in-World: World Models in a Closed-Loop World" (ICLR'26 Oral)
iChristGit
Custom tools to enhance your Open-Webui Experience! 🚀
mantoufan
创意→完整短剧剧本→Seedance 2.0 视频提示词 全链路 Claude Skill:上游写 50-100 集短剧剧本,下游转标准剧本/资产/分镜;含 12 铁律 + 17 模板 + 视听语言·去AI感·人物真实感技巧库。 | Full idea→screenplay→Seedance 2.0 video-prompt pipel…
Echo-Team-Joy-Future-Academy-JD
A Simple Baseline for Video World Models with Memory
guaardvark
The self-hosted AI workstation. Autonomous screen agents, 3-tier neural routing, parallel agent swarms, video generation, 4K/8K upscaling, RAG, voice interface, 70+ tool…
thevickypedia
Fully Functional Voice Based Natural Language UI
denizsafak
Generate audiobooks from EPUBs, PDFs and text with synchronized captions.
ARahim3
Fine-tune LLMs on your Mac with Apple Silicon. SFT, DPO, GRPO, Vision, TTS, STT, Embedding, and OCR fine-tuning — natively on MLX. Unsloth-compatible API.
transcriptionstream
turnkey self-hosted offline transcription and diarization service with llm summary
saharmor
Build real time speech2text web apps using OpenAI's Whisper https://openai.com/blog/whisper/
nonwill
GoldenDict++: Optimizations for faster dictionary loading and searching, even with large dictionary collections. OCR integration for text recognition, enhanced media pla…
zarazhangrui
Turn any content into a personalized AI podcast. NotebookLM-style, except you control the script, voices, and hosts. Listen in Apple Podcasts, Spotify, or any podcast ap…
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
babysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time