No visual example yet
Explore the skillMlx Audio
Blaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
1–16 / 64
Results: 64
No visual example yet
Explore the skillBlaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
No visual example yet
Explore the skillpytorch
Data manipulation and transformation for audio signal processing, powered by PyTorch
No visual example yet
Explore the skilllibAudioFlux
A library for audio and music analysis, feature extraction.
No visual example yet
Explore the skilltyiannak
Python Audio Analysis Library: Feature Extraction, Classification, Segmentation and Applications
No visual example yet
Explore the skilliver56
A Python library for audio data augmentation. Useful for making audio ML models work well in the real world, not just in the lab.
No visual example yet
Explore the skillhkchengrex
[CVPR 2025] MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis
No visual example yet
Explore the skilldiodiogod
A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwen3-TTS, Cozy Voice 3, Step Audio…
No visual example yet
Explore the skillarchinetai
A timeline of the latest AI models for audio generation, starting in 2023!
No visual example yet
Explore the skillgitmylo
A webui for different audio related Neural Networks
No visual example yet
Explore the skillBinWang28
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
No visual example yet
Explore the skillBlaizzy
A modular Swift SDK for audio processing with MLX on Apple Silicon
No visual example yet
Explore the skillstepfun-ai
A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics, and features robust zero-shot…
No visual example yet
Explore the skillGoogleChrome
Audion is a Chrome extension that adds a Web Audio panel to Developer Tools. This panel visualizes the web audio graph in real-time.
No visual example yet
Explore the skillFunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
No visual example yet
Explore the skillnexu-io
Audio generation skill — jingles, beds, voiceover, and sound effects. Routes music requests to Suno V5 / Udio / Lyria, speech to MiniMax TTS / FishAudio / ElevenLabs V3,…
No visual example yet
Explore the skillcalesthio
Generate expressive, multilingual narration with fish.audio (S1 / S2-generation models) and reuse cloned voices via reference_id. Use when the user prefers fish.audio/Fi…