Whisper Diarization
MahmoudAshraf97
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
33–48 / 82
Results: 82
MahmoudAshraf97
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
flashlight
Facebook AI Research's Automatic Speech Recognition Toolkit
mkiol
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
titanwings
将同事、导师、搭档的工作经验和性格永久保存为 AI Skill。提供飞书、钉钉、Slack、微信聊天记录、邮件等多源数据采集,生成真正能替他工作的 AI Skill——用他的技术规范写代码,用他的语气回答问题,知道他什么时候会甩锅。
wechat-article
微信公众号文章批量下载工具,支持导出阅读量与评论数据。无需搭建环境,支持在线使用、Docker 私有化部署和 Cloudflare 部署。支持多种格式导出,HTML 格式可100%还原文章排版与样式。
dontbesilent2025
从 12,307 条推文中提炼的商业诊断方法论,做成 Claude Code skill。包含商业模式诊断、对标分析、内容创作诊断、执行力诊断、概念拆解 5 个工具。附带 4,176 个结构化知识原子,可用于 RAG 知识库。
modelscope
Open-source, accurate and easy-to-use video speech recognition & clipping tool. LLM-based AI clipping integrated.
babysor
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
OpenBMB
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
index-tts
An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
k2-fsa
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Su…
coqui-ai
🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
snakers4
Silero Models: pre-trained text-to-speech models made embarrassingly simple
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.