No visual example yet
Explore the skillFireRedASR
FireRedTeam
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering ou…
OPENAGENTSKILL / DIRECTORY
Temukan skill untuk tugas berikutnya dengan Codex, Claude Code, Cursor, dan lainnya.
28 Skills
Hasil: 28
No visual example yet
Explore the skillFireRedTeam
Open-source industrial-grade ASR models supporting Mandarin, Chinese dialects and English, achieving a new SOTA on public Mandarin ASR benchmarks, while also offering ou…
No visual example yet
Explore the skillFunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
No visual example yet
Explore the skillahmetoner
OpenAI Whisper ASR Webservice API
No visual example yet
Explore the skillFireRedTeam
A SOTA Industrial-Grade All-in-One ASR system with ASR, VAD, LID, and Punc modules. FireRedASR2 supports Chinese (Mandarin, 20+ dialects/accents), English, code-switchin…
No visual example yet
Explore the skillkaldi-asr
kaldi-asr/kaldi is the official location of the Kaldi project.
No visual example yet
Explore the skillFunAudioLLM
Multilingual speech understanding: ASR + emotion recognition + audio event detection. 50+ languages, 15x faster than Whisper, non-autoregressive.
No visual example yet
Explore the skillPaddlePaddle
Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…
No visual example yet
Explore the skillPeterH0323
Streamer-Sales 销冠 —— 卖货主播 LLM 大模型🛒🎁,一个能够根据给定的商品特点从激发用户购买意愿角度出发进行商品解说的卖货主播大模型。🚀⭐内含详细的数据生成流程❗ 📦另外还集成了 LMDeploy 加速推理🚀、RAG检索增强生成 📚、TTS文字转语音🔊、数字人生成 🦸、 Agent 使用网络查询实时信…
No visual example yet
Explore the skillAutoArk
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
No visual example yet
Explore the skillsoniqo
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
No visual example yet
Explore the skillAkagawaTsurunaki
AI VTuber with LLM, ASR, TTS, OCR, CV and more technologies to live stream or play Minecraft with you.
No visual example yet
Explore the skillcalesthio
DashScope (Alibaba Cloud Bailian / 阿里云百炼) integration — image generation (qwen-image-2.0-pro), text-to-speech (qwen3-tts-flash), and ASR with word-level timestamps (qwen…
No visual example yet
Explore the skillelevenyellow
An AI-powered interactive avatar engine using Live2D, LLM, ASR, TTS, and RVC. Ideal for VTubing, streaming, and virtual assistant applications.
No visual example yet
Explore the skillcoqui-ai
💎 A list of accessible speech corpora for ASR, TTS, and other Speech Technologies
No visual example yet
Explore the skillPluviobyte
视频内容逐字稿提取专用 skill。Use when the user provides a 抖音 or 小红书 URL and asks for逐字稿、视频转文字、听写视频或提取口播内容。Only extract transcript content; do not use this skill as the production s…
No visual example yet
Explore the skillfluxions-ai
Real-time voice assistant — WebRTC streaming, faster-whisper ASR, local LLM, Vui Nano (300M) TTS. OpenAI Realtime API compatible. Voice cloning, barge-in, ~9× realtime o…