Fun ASR
FunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
65–80 / 158
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 158
FunAudioLLM
End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.
Pluviobyte
Create original editorial technology covers with a controlled layout, custom diagram generation, and editable SVG or PNG output.
jacobgil
Advanced AI Explainability for computer vision. Support for CNNs, Vision Transformers, Classification, Object detection, Segmentation, Image similarity and more.
libvips
A fast image processing library with low memory needs.
ConardLi
ConardLi's open-source Skills collection, featuring web design, knowledge retrieval, image generation, and more.
hwdsl2
Docker image to run an IPsec VPN server, with IPsec/L2TP, Cisco IPsec and IKEv2. Auto-generates server config and supports VPN client setup on Linux, Windows, macOS, iOS…
ningzimu
GPT-Image-2 PPT Generator Skill for Creating Image-Based PowerPoint Presentations in Codex and Other Skill-Compatible Agents
zuruoke
a machine learning image inpainting task that instinctively removes watermarks from image indistinguishable from the ground truth image
nexu-io
🎨 Local-first, open-source Claude Design alternative. 🖥️ Native desktop app. ⚡ 259+ Skills · ✨ 142+ Design Systems 🖼️ Web · desktop · mobile prototypes · slides · ima…
elegantapp
Automates PWA asset generation and image declaration. Automatically generates icon and splash screen images, favicons and mstile images. Updates manifest.json and index.…
voxelmorph
Unsupervised Learning for Image Registration
CVCUDA
CV-CUDA™ is an open-source, GPU accelerated library for cloud-scale image processing and computer vision.
OpenStitching
A Python package for fast and robust Image Stitching
YouMind-OpenLab
AI skill for OpenClaw & Claude Code — recommend from 10000+ Nano Banana Pro (Gemini) image prompts. Smart search by use case, content remix, sample images.
RVC-Boss
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)