No visual example yet
Explore the skillMlx Audio
Blaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
OPENAGENTSKILL / DIRECTORY
次のタスクに合うスキルを。Codex、Claude Code、Cursor などのツールを探せます。
426 Skills
検索結果: 426
No visual example yet
Explore the skillBlaizzy
A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
No visual example yet
Explore the skillailia-ai
The collection of pre-trained, state-of-the-art AI models for ailia SDK
No visual example yet
Explore the skillspotify
No visual example yet
Explore the skillmodelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
No visual example yet
Explore the skillNVIDIA
A GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learning training and inference appl…
No visual example yet
Explore the skillEventual-Inc
High-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale
No visual example yet
Explore the skillvllm-project
A framework for efficient model inference with omni-modality models
No visual example yet
Explore the skillwhitphx
Real-time video and audio processing on Streamlit
No visual example yet
Explore the skillWyattBlue
No visual example yet
Explore the skillAgents365-ai
Automated 4K video podcast creation for coding agents
No visual example yet
Explore the skillOpenMOSS
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and…
No visual example yet
Explore the skillpytorch
Data manipulation and transformation for audio signal processing, powered by PyTorch
No visual example yet
Explore the skillopen-mmlab
Amphion (/æmˈfaɪən/) is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get…
No visual example yet
Explore the skillmiantiao-me
一个基于 AI 的 Hacker News 中文播客项目,每天自动抓取 Hacker News 热门文章,通过 AI 生成中文总结并转换为播客内容。
No visual example yet
Explore the skillzebbern
80+ free AI services for chat, image, video, voice & APIs (may sometimes include access to lead gen ai models for free)
No visual example yet
Explore the skillpnnbao97
Vietnamese TTS with instant voice cloning • On-device • Real-time CPU inference • 24kHz audio quality • Chuyển văn bản thành giọng nói tiếng Việt • Text to speech tiếng…