No visual example yet
Explore the skillConformer
sooftware
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
Find a skill for your next task. Preview examples where available.
49–64 / 94
Results: 94
No visual example yet
Explore the skillsooftware
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
No visual example yet
Explore the skillsemperai
Amica is an open source interface for interactive communication with 3D characters with voice synthesis and speech recognition.
No visual example yet
Explore the skillalphacep
Offline speech recognition for Android with Vosk library.
No visual example yet
Explore the skillzubair-trabzada
Schema.org structured data audit and generation optimized for AI discoverability — detect, validate, and generate JSON-LD markup
No visual example yet
Explore the skillictnlp
StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.
No visual example yet
Explore the skillsdkcarlos
A voice control - voice commands - speech recognition and speech synthesis javascript library. Create your own siri,google now or cortana with Google Chrome within your…
No visual example yet
Explore the skillalphacep
WebSocket, gRPC and WebRTC speech recognition server based on Vosk and Kaldi libraries
No visual example yet
Explore the skillopensemanticsearch
Open Source research tool to search, browse, analyze and explore large document collections by Semantic Search Engine and Open Source Text Mining & Text Analytics platfo…
No visual example yet
Explore the skillalumae
Real-time full-duplex speech recognition server, based on the Kaldi toolkit and the GStreamer framwork.
No visual example yet
Explore the skillemedvedev
A Tensorflow model for text recognition (CNN + seq2seq with visual attention) available as a Python package and compatible with Google Cloud ML Engine.
No visual example yet
Explore the skillsavbell
💬📝 A small dictation app using OpenAI's Whisper speech recognition model.
No visual example yet
Explore the skillhukenovs
HAnd Gesture Recognition Image Dataset
No visual example yet
Explore the skillTensorSpeech
:zap: TensorFlowASR: Almost State-of-the-art Automatic Speech Recognition in Tensorflow 2. Supported languages that can use characters or subwords
No visual example yet
Explore the skillsamc621
All-in-one bot, with auto captcha-solving and proxy management, using Node.js and Puppeteer.
No visual example yet
Explore the skillBinWang28
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
No visual example yet
Explore the skillTikHub
High-performance asynchronous Douyin(抖音) TikTok Xiaohongshu(小红书) Kuaishou(快手) Weibo(微博) Instagram YouTube(油管) Twitter(X) Captcha Solver(验证码解决器) Temp Mail(临时邮箱) API(接口).