Instant voice cloning by MIT and MyShell. Audio foundation model.
$ npx skills add myshell-ai/OpenVoiceScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
AI Agent Skill Repository
Browse reusable skills for Codex, Claude Code, Cursor, finance, research, web scraping, PPT, football analytics, data, marketing, design, and more.
Live registry search
Exact name and slug matches are checked against the live registry before ranked alternatives.
Decision filters
Showing 1-16 of 109 ranked candidates matching "offline-voice"
Best blend of relevance, quality, freshness, and verified outcomes
Instant voice cloning by MIT and MyShell. Audio foundation model.
$ npx skills add myshell-ai/OpenVoiceScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine
$ npx skills add netease-youdao/EmotiVoiceScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your C…
$ npx skills add Enemyx-net/VibeVoice-ComfyUIScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channe…
$ npx skills add rapidaai/voice-aiScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and multi-speaker dialogue
$ npx skills add Saganaki22/ComfyUI-OmniVoice-TTSScenario Design and creative · I need my agent to produce design assets, UI directions, presentations, or creative media workflows.
CLI + Codex · 4 targets
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
$ npx skills add FunAudioLLM/CosyVoiceScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
OpenAI Agents + CLI · 4 targets
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTu…
$ npx skills add abus-aikorea/voice-proScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
The open-source ElevenLabs alternative for local voice cloning, design, create, dubbing and dictation Desktop App
$ npx skills add debpalash/OmniVoice-StudioScenario Design and creative · I need my agent to produce design assets, UI directions, presentations, or creative media workflows.
CLI + Codex · 4 targets
Voice AI SDK is a reusable Android library that gives any app a full voice-driven AI conversation pipeline in minutes. Voice Assistant + Android Voide AI + SDK + MVVM +…
$ npx skills add ahmedeltaher/Android-MVVM-Architecture-Android-Voice-AI-SDKScenario Coding agents · I need a coding agent that can understand a repository, edit code, and review pull requests.
CLI + Codex · 4 targets
Offline multi-agent simulation & prediction engine. English fork of MiroFish with Neo4j + Ollama local stack.
$ npx skills add nikmcfly/MiroFish-OfflineScenario Coding agents · I need a coding agent that can understand a repository, edit code, and review pull requests.
CLI + Codex · 4 targets
Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.
$ npx skills add janvarev/Irene-Voice-AssistantScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)
$ npx skills add VRCWizard/TTS-Voice-WizardScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
🗣 An overlay that gets your user’s voice permission and input as text in a customizable UI
$ npx skills add algolia/voice-overlay-iosScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
📱 Browser add-on allowing you to quickly generate a QR code offline with the URL of the open tab or other text!
$ npx skills add rugk/offline-qr-codeScenario Legal and compliance · I need my agent to review contracts, privacy policies, or compliance documents and summarize risks.
Browser agents + CLI · 4 targets
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
$ npx skills add mkiol/dsnoteScenario Design and creative · I need my agent to produce design assets, UI directions, presentations, or creative media workflows.
CLI + Codex · 4 targets
Offline Text To Speech synthesis for python
$ npx skills add nateshmbhat/pyttsx3Scenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
Page 1
Showing the strongest 16 results to keep the registry fast for humans and agents. Refine by use case, platform, stars, or search query for a narrower shortlist.
Try the agent resolve APIAI Agent Skill Repository
Browse reusable skills for Codex, Claude Code, Cursor, finance, research, web scraping, PPT, football analytics, data, marketing, design, and more.
Live registry search
Exact name and slug matches are checked against the live registry before ranked alternatives.
Decision filters
Showing 1-16 of 109 ranked candidates matching "offline-voice"
Best blend of relevance, quality, freshness, and verified outcomes
Instant voice cloning by MIT and MyShell. Audio foundation model.
$ npx skills add myshell-ai/OpenVoiceScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
EmotiVoice 😊: a Multi-Voice and Prompt-Controlled TTS Engine
$ npx skills add netease-youdao/EmotiVoiceScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice synthesis directly within your C…
$ npx skills add Enemyx-net/VibeVoice-ComfyUIScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channe…
$ npx skills add rapidaai/voice-aiScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and multi-speaker dialogue
$ npx skills add Saganaki22/ComfyUI-OmniVoice-TTSScenario Design and creative · I need my agent to produce design assets, UI directions, presentations, or creative media workflows.
CLI + Codex · 4 targets
Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.
$ npx skills add FunAudioLLM/CosyVoiceScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
OpenAI Agents + CLI · 4 targets
Gradio WebUI for creators and developers, featuring key TTS (Edge-TTS, kokoro) and zero-shot Voice Cloning (E2 & F5-TTS, CosyVoice), with Whisper audio processing, YouTu…
$ npx skills add abus-aikorea/voice-proScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
The open-source ElevenLabs alternative for local voice cloning, design, create, dubbing and dictation Desktop App
$ npx skills add debpalash/OmniVoice-StudioScenario Design and creative · I need my agent to produce design assets, UI directions, presentations, or creative media workflows.
CLI + Codex · 4 targets
Voice AI SDK is a reusable Android library that gives any app a full voice-driven AI conversation pipeline in minutes. Voice Assistant + Android Voide AI + SDK + MVVM +…
$ npx skills add ahmedeltaher/Android-MVVM-Architecture-Android-Voice-AI-SDKScenario Coding agents · I need a coding agent that can understand a repository, edit code, and review pull requests.
CLI + Codex · 4 targets
Offline multi-agent simulation & prediction engine. English fork of MiroFish with Neo4j + Ollama local stack.
$ npx skills add nikmcfly/MiroFish-OfflineScenario Coding agents · I need a coding agent that can understand a repository, edit code, and review pull requests.
CLI + Codex · 4 targets
Ирина - русский голосовой ассистент для работы оффлайн. Поддерживает скиллы через плагины.
$ npx skills add janvarev/Irene-Voice-AssistantScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
Speech to Text to Speech. Song now playing. Sends text as OSC messages to VRChat to display on avatar. (STTTS) (Speech to TTS) (VRC STT System) (VTuber TTS)
$ npx skills add VRCWizard/TTS-Voice-WizardScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
🗣 An overlay that gets your user’s voice permission and input as text in a customizable UI
$ npx skills add algolia/voice-overlay-iosScenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
📱 Browser add-on allowing you to quickly generate a QR code offline with the URL of the open tab or other text!
$ npx skills add rugk/offline-qr-codeScenario Legal and compliance · I need my agent to review contracts, privacy policies, or compliance documents and summarize risks.
Browser agents + CLI · 4 targets
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
$ npx skills add mkiol/dsnoteScenario Design and creative · I need my agent to produce design assets, UI directions, presentations, or creative media workflows.
CLI + Codex · 4 targets
Offline Text To Speech synthesis for python
$ npx skills add nateshmbhat/pyttsx3Scenario Multimodal media · I need my agent to process images, video, or audio and extract useful information.
CLI + Codex · 4 targets
Page 1
Showing the strongest 16 results to keep the registry fast for humans and agents. Refine by use case, platform, stars, or search query for a narrower shortlist.
Try the agent resolve API