No visual example yet
Explore the skillWav2letter
flashlight
Facebook AI Research's Automatic Speech Recognition Toolkit
OPENAGENTSKILL / DIRECTORY
Trouvez un skill pour votre prochaine tâche avec Codex, Claude Code, Cursor et plus encore.
107 Skills
Résultats: 107
No visual example yet
Explore the skillflashlight
Facebook AI Research's Automatic Speech Recognition Toolkit
No visual example yet
Explore the skilljianchang512
Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式
No visual example yet
Explore the skillmkiol
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
No visual example yet
Explore the skillhuggingface
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
No visual example yet
Explore the skillFireRedTeam
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering ou…
No visual example yet
Explore the skillFunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
No visual example yet
Explore the skillroyshil
OBS plugin for local speech recognition and captioning using AI
No visual example yet
Explore the skillk2-fsa
Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows,…
No visual example yet
Explore the skillyeyupiaoling
Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inf…
No visual example yet
Explore the skillGauravSingh9356
Personal Assistant built using python libraries. It does almost anything which includes sending emails, Optical Text Recognition, Dynamic News Reporting at any time with…
No visual example yet
Explore the skillsooftware
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
No visual example yet
Explore the skillalphacep
Offline speech recognition for Android with Vosk library.
No visual example yet
Explore the skillk2-fsa
Speech-to-text server framework with next-gen Kaldi
No visual example yet
Explore the skillBinWang28
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
No visual example yet
Explore the skillsandrohanea
Whisper.net. Speech to text made simple using Whisper Models
No visual example yet
Explore the skillageitgey
The world's simplest facial recognition api for Python and the command line