Offline nonclinical phone-workflow demonstration over supplied speech/pause durations. Returns illustrative pause-ratio labels, not medical findings or an actual emergen…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
Page 36 · 16 shown · 743 public entries
Results: 743
Offline nonclinical phone-workflow demonstration over supplied speech/pause durations. Returns illustrative pause-ratio labels, not medical findings or an actual emergen…
travisjneuman
Professional audio production for music, podcasts, and sound design. Use when working with audio recording, mixing, mastering, or sound design for any medium.
unknowlei
Default downstream MiniMax H3 specialist for every image-based request unless the user explicitly declares boundary-only first/last frames with no reusable reference rol…
unknowlei
Downstream MiniMax H3 specialist for professional text-to-video (T2VA) prompts using the official three-field format. Use after minimax-h3-creative-director routes a req…
sbarex
MacOS Finder Extension to show information about media files (images, video and audio), PDF and Office files on the contextual menu.
Rongjiehuang
PyTorch Implementation of ProDiff (ACM-MM'22) with a Extremely-Fast diffusion speech synthesis pipeline
alan-ai
The Self-Coding System for Your App — Alan AI SDK for Power Apps
Rongjiehuang
PyTorch Implementation of FastDiff (IJCAI'22)
hassancs91
Generates one expressive narration MP3 per scene for an English storybook, using the ElevenLabs MCP (text_to_speech). Reads {slug}_scenes.json (from scene-splitter), pro…
rafaballerini
Assistente pessoal virtual desenvolvida com Python 🤖
ivanvovk
Implementation of WaveGrad high-fidelity vocoder from Google Brain in PyTorch.
TornadoInsight
OpenAi-Sora (SoraFlows) is an open-source, cross-platform web application for AI-powered video creation and editing using the latest OpenAI Sora model. Effortlessly gene…
hassancs91
Orchestrator that runs the full AI Storybook pipeline end-to-end on one English story. Dispatches the 4 component skills in order — scene-splitter → story-illustrator →…
r9y9
Library to build speech synthesis systems designed for easy and fast prototyping.
2b-t
Guide on how to set-up Linux and Docker for real-time applications using the Ubuntu realtime-kernel/PREEMPT_RT patch with a focus on robotics with ROS and ROS 2
ictnlp
Stream-Omni is a GPT-4o-like language-vision-speech chatbot that simultaneously supports interaction across various modality combinations.