No visual example yet
Explore the skillFun ASR
FunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
OPENAGENTSKILL / DIRECTORY
Finde den passenden Skill für deine nächste Aufgabe mit Codex, Claude Code, Cursor und mehr.
13 Skills
Ergebnisse: 13
No visual example yet
Explore the skillFunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
No visual example yet
Explore the skillFunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
No visual example yet
Explore the skillk2-fsa
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Su…
No visual example yet
Explore the skillespeak-ng
eSpeak NG is an open source speech synthesizer that supports more than hundred languages and accents.
No visual example yet
Explore the skillRHVoice
a free and open source speech synthesizer for Russian and other languages
No visual example yet
Explore the skillTensorSpeech
:stuck_out_tongue_closed_eyes: TensorFlowTTS: Real-Time State-of-the-art Speech Synthesis for Tensorflow 2 (supported including English, French, Korean, Chinese, German…
No visual example yet
Explore the skillnazdridoy
A CLI text-to-speech tool using the Kokoro model, supporting multiple languages, voices (with blending), and various input formats including EPUB books and PDF documents.
No visual example yet
Explore the skillDigitalPhonetics
Controllable and fast Text-to-Speech for over 7000 languages!
No visual example yet
Explore the skillAzure-Samples
Microsoft Text-to-Speech API sample code in several languages, part of Cognitive Services.
No visual example yet
Explore the skillTensorSpeech
:zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords
No visual example yet
Explore the skillpnlpal
📚 A customizable dictionary extension that supports double-click lookups in 20+ languages, 1000+ dictionaries, text-to-speech, translation and Anki integration.
No visual example yet
Explore the skillNotely-Voice
A 100% private AI voice transcription app that converts speech to text in 100+ languages. Built with Compose Multiplatform for Android & iOS using Whisper AI - no cloud…
No visual example yet
Explore the skillSamirPaulb
A desktop application that uses AI to translate voice between languages in real time, while preserving the speaker's tone and emotion.