No visual example yet
Explore the skillWhisper Diarization
MahmoudAshraf97
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
OPENAGENTSKILL / DIRECTORY
次のタスクに合うスキルを。Codex、Claude Code、Cursor などのツールを探せます。
11 Skills
検索結果: 11
No visual example yet
Explore the skillMahmoudAshraf97
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
No visual example yet
Explore the skillm-bain
WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)
No visual example yet
Explore the skillhomelab-00
A fully local and private Speech-To-Text app with cross-platform support, speaker diarization, Audio Notebook mode, LM Studio integration, and both longform and live tra…
No visual example yet
Explore the skilltranscriptionstream
turnkey self-hosted offline transcription and diarization service with llm summary
No visual example yet
Explore the skillizwi-ai
Voice AI runtime. Local first transcription, speaker diarization, TTS, and voice cloning with an OpenAI compatible API.
No visual example yet
Explore the skillk2-fsa
Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Su…
No visual example yet
Explore the skillFunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
No visual example yet
Explore the skillsoniqo
AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML
No visual example yet
Explore the skilljianshuo
Use when the user has a video + a target-language SRT and wants the video to actually speak that language — generates a time-aligned TTS voice dub. Routes by voice ID —…
No visual example yet
Explore the skillTranscribe meeting audio with speaker diarization, generate structured summaries with action items, decisions, and follow-ups, and support multiple audio formats and lan…
No visual example yet
Explore the skillagents-inc
Speech-to-text transcription and translation via OpenAI Audio API -- models, response formats, timestamps, prompting, streaming, chunking, and diarization