No visual example yet
Explore the skillIrodori TTS
Aratako
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
81–96 / 103
Results: 103
No visual example yet
Explore the skillAratako
A Flow Matching-based Text-to-Speech Model with Emoji-driven Style Control
No visual example yet
Explore the skillVonage
Vonage REST API client for PHP. API support for SMS, Voice, Text-to-Speech, Numbers, Verify (2FA) and more.
No visual example yet
Explore the skillKieirra
Fully local, private and cross platform Speech-to-Text with LLM Post-processing
No visual example yet
Explore the skillwatzon
A native macOS menu bar dictation app using local speech-to-text with WhisperKit
No visual example yet
Explore the skillcboard-org
Augmentative and Alternative Communication (AAC) system with text-to-speech for the browser
No visual example yet
Explore the skillpnlpal
📚 A customizable dictionary extension that supports double-click lookups in 20+ languages, 1000+ dictionaries, text-to-speech, translation and Anki integration.
No visual example yet
Explore the skill2noise
A generative speech model for daily dialogue.
No visual example yet
Explore the skillxorbitsai
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all throug…
No visual example yet
Explore the skillmodelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
No visual example yet
Explore the skillinstillai
:speech_balloon: Machine Learning Course with Python:
No visual example yet
Explore the skilllinto-ai
Multilingual Automatic Speech Recognition with word-level timestamps and confidence
No visual example yet
Explore the skillkubeai-project
AI Inference Operator for Kubernetes. The easiest way to serve ML models in production. Supports VLMs, LLMs, embeddings, and speech-to-text.
No visual example yet
Explore the skillnexu-io
Audio generation skill — jingles, beds, voiceover, and sound effects. Routes music requests to Suno V5 / Udio / Lyria, speech to MiniMax TTS / FishAudio / ElevenLabs V3,…
No visual example yet
Explore the skillnazdridoy
A CLI text-to-speech tool using the Kokoro model, supporting multiple languages, voices (with blending), and various input formats including EPUB books and PDF documents.
No visual example yet
Explore the skillcalesthio
DashScope (Alibaba Cloud Bailian / 阿里云百炼) integration — image generation (qwen-image-2.0-pro), text-to-speech (qwen3-tts-flash), and ASR with word-level timestamps (qwen…
No visual example yet
Explore the skillcalesthio
Generate Mandarin and multilingual narration with Volcengine Doubao Speech 2.0. Use when creating Chinese voiceovers, when the user prefers Doubao/Volcengine/火山引擎/豆包 TTS…