No visual example yet
Explore the skillKaldi Gstreamer Server
alumae
Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
33–41 / 41
Results: 41
No visual example yet
Explore the skillalumae
Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.
No visual example yet
Explore the skillsavbell
💬📝 A small dictation app using OpenAI's Whisper speech recognition model.
No visual example yet
Explore the skillTensorSpeech
:zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords
No visual example yet
Explore the skillBinWang28
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
No visual example yet
Explore the skillmodelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
No visual example yet
Explore the skilllinto-ai
Multilingual Automatic Speech Recognition with word-level timestamps and confidence
No visual example yet
Explore the skillbyjlw
Analyze videos using LLMs, Computer Vision and Automatic Speech Recognition
No visual example yet
Explore the skillmaxazure
OpenClaw/Codex Skill: Auto video editing for talk/vlog videos — speech recognition, sentence splitting, subtitle burning, and clip merging
No visual example yet
Explore the skillnonwill
GoldenDict++: Optimizations for faster dictionary loading and searching, even with large dictionary collections. OCR integration for text recognition, enhanced media pla…