No visual example yet
Explore the skillWhisperX
m-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
97 Skills
Results: 97
No visual example yet
Explore the skillm-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
No visual example yet
Explore the skillFunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
No visual example yet
Explore the skillserengil
A Lightweight Face Recognition and Facial Attribute Analysis (Age, Gender, Emotion and Race) Library for Python
No visual example yet
Explore the skillalphacep
Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node
No visual example yet
Explore the skillTalAter
💬 Speech recognition for your site
No visual example yet
Explore the skillmindee
docTR (Document Text Recognition) - a seamless, high-performing & accessible library for OCR-related tasks powered by Deep Learning.
No visual example yet
Explore the skillwenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
No visual example yet
Explore the skillMahmoudAshraf97
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
No visual example yet
Explore the skillflashlight
Facebook AI Research's Automatic Speech Recognition Toolkit
No visual example yet
Explore the skilljianchang512
Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式
No visual example yet
Explore the skillThoughtfulDev
Stalk your Friends. Find their Instagram, FB and Twitter Profiles using Image Recognition and Reverse Image Search.
No visual example yet
Explore the skillhuggingface
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
No visual example yet
Explore the skillkenshohara
3D ResNets for Action Recognition (CVPR 2018)
No visual example yet
Explore the skillFireRedTeam
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering ou…
No visual example yet
Explore the skillFunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
No visual example yet
Explore the skillroyshil
OBS plugin for local speech recognition and captioning using AI