OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
33–48 / 177
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 177
NVIDIA
State-of-the-Art Deep Learning scripts organized by models - easy to train and deploy with reproducible accuracy and performance on enterprise-grade infrastructure.
wenet-e2e
Production First and Production Ready End-to-End Speech Recognition Toolkit
Picovoice
On-device wake word detection powered by deep learning
MahmoudAshraf97
Automatic Speech Recognition with Speaker Diarization based on OpenAI Whisper
flashlight
Facebook AI Research's Automatic Speech Recognition Toolkit
jianchang512
Voice Recognition to Text Tool / 一个离线运行的本地音视频转字幕工具,输出json、srt字幕、纯文字格式
AlekPet
Custom nodes that extend the capabilities of Comfyui
fikrikarim
On-device, real-time multimodal AI. Have natural voice and vision conversations with an AI that runs entirely on your machine. Powered by Gemma 4 E2B and Kokoro.
sanchit-gandhi
JAX implementation of OpenAI's Whisper model for up to 70x speed-up on TPU.
mkiol
Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and Machine translation.
Purfview
Whisper & Faster-Whisper standalone executables for those who don't want to bother with Python.
bytedance
SALMONN family: A suite of advanced multi-modal LLMs
huggingface
Distilled variant of Whisper for speech recognition. 6x faster, 50% smaller, within 1% word error rate.
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.