No visual example yet
Explore the skillFireRedASR
FireRedTeam
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering ou…
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
38 Skills
Results: 38
No visual example yet
Explore the skillFireRedTeam
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering ou…
No visual example yet
Explore the skillFunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
No visual example yet
Explore the skillroyshil
OBS plugin for local speech recognition and captioning using AI
No visual example yet
Explore the skillk2-fsa
Real-time speech recognition and voice activity detection (VAD) using next-gen Kaldi with ncnn without Internet connection. Support iOS, Android, Linux, macOS, Windows,…
No visual example yet
Explore the skillyeyupiaoling
Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inf…
No visual example yet
Explore the skillmravanelli
pytorch-kaldi is a project for developing state-of-the-art DNN/RNN hybrid speech recognition systems. The DNN part is managed by pytorch, while feature extraction, label…
No visual example yet
Explore the skillnobody132
中文语音识别; Mandarin Automatic Speech Recognition;
No visual example yet
Explore the skilljulius-speech
Open-Source Large Vocabulary Continuous Speech Recognition Engine
No visual example yet
Explore the skillastorfi
:unlock: Lip Reading - Cross Audio-Visual Recognition using 3D Architectures
No visual example yet
Explore the skillsyhw
Attempt at tracking states of the arts and recent results (bibliography) on speech recognition.
No visual example yet
Explore the skillGauravSingh9356
Personal Assistant built using python libraries. It does almost anything which includes sending emails, Optical Text Recognition, Dynamic News Reporting at any time with…
No visual example yet
Explore the skillsooftware
[Unofficial] PyTorch implementation of "Conformer: Convolution-augmented Transformer for Speech Recognition" (INTERSPEECH 2020)
No visual example yet
Explore the skillalphacep
Offline speech recognition for Android with Vosk library.
No visual example yet
Explore the skillBinWang28
The hub for audio AI research: papers, open models, benchmarks & datasets across audio LLMs, speech recognition, TTS, music & audio generation.
No visual example yet
Explore the skillmodelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
No visual example yet
Explore the skillcalesthio
Create AI avatar videos with precise control over avatars, voices, scripts, scenes, and backgrounds using HeyGen's v2 API. Use when: (1) Choosing a specific avatar and v…