AI Agent Skill Repository

AI Agent Skills Directory

Browse reusable skills for Codex, Claude Code, Cursor, finance, research, web scraping, PPT, football analytics, data, marketing, design, and more.

Browse by scenario

Real skills, grouped by the work your agent needs to finish.

Each directory entry links to real skill pages with GitHub adoption, trust score, install handoff, risk notes, and agent-readable metadata. Use these as starting points when you want a shortlist before asking an agent to install anything.

Design to deployment

Frontend and UI skills

Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.

  • Inference9.3K stars

    Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and mult…

    Trust 87Quality 100ml-automation
  • Kokoro FastAPI5.1K stars

    Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model w/multiplatform CPU, AMD, NVIDIA GPU PyTorch s…

    Trust 87Quality 99media-automation
  • Dsnote1.5K stars

    Speech Note Linux app. Note taking, reading and translating with offline Speech to Text, Text to Speech and M…

    Trust 87Quality 94media-automation

Codex, Claude Code, Cursor

Coding agent skills

Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.

  • Skills122 stars

    Agent Skills for the Venice.ai API. One folder per surface area, each with a SKILL.md for agent runtimes (Cur…

    Trust 84Quality 87utility
  • 「說人話」:繁體中文的去 AI 味改寫 skill。抓 38 種 AI 寫作痕跡,順手校正中國用語與半形標點,給 Claude Code / Codex / Cursor 用。

    Trust 87Quality 95utility
  • 将同事、导师、搭档的工作经验和性格永久保存为 AI Skill。提供飞书、钉钉、Slack、微信聊天记录、邮件等多源数据采集,生成真正能替他工作的 AI Skill——用他的技术规范写代码,用他的语气回答问题,知道他什…

    Trust 84Quality 100chinese

Documents and knowledge

Research and RAG skills

Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.

  • Markpdfdown1.8K stars

    A high-quality PDF to Markdown tool based on large language model visual recognition. 一款基于大模型视觉识别的高质量PDF转Mark…

    Trust 83Quality 84document-processing
  • EasyOCR30K stars

    Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabi…

    Trust 87Quality 94document-processing
  • Chinese copywriting guidelines for better written communication/中文文案排版指北

    Trust 82Quality 78document-processing

Markets and quant

Finance and trading skills

Stock analysis, market research, quant backtesting, financial data, and investment research skills.

  • Quant Trading10K stars

    Python quantitative trading strategies including VIX Calculator, Pattern Recognition, Commodity Trading Advis…

    Trust 82Quality 82finance

Crawlers and extraction

Web scraping skills

Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.

  • 微信公众号文章批量下载工具,支持导出阅读量与评论数据。无需搭建环境,支持在线使用、Docker 私有化部署和 Cloudflare 部署。支持多种格式导出,HTML 格式可100%还原文章排版与样式。

    Trust 85Quality 100chinese
  • Lue787 stars

    Terminal eBook Reader with Audiobook-Quality Text-to-Speech — Supports EPUB, PDF, DOCX, HTML, RTF, TXT, and M…

    Trust 78Quality 80document-processing
  • Pot Desktop19K stars

    🌈一个跨平台的划词翻译和OCR软件 | A cross-platform software for text translation and recognition.

    Trust 90Quality 100document-processing

Slides and decks

PPT and presentation skills

Presentation generation, editable PPTX, slide decks, speaker notes, and visual storytelling workflows.

  • Kill Ai Slop786 stars

    A field guide to the visual & copy tics of AI-generated products — and an Agent Skill that scans your project…

    Trust 87Quality 95design-creative

Prompts, B-roll, explainers

Video creation skills

Video-generation prompts, B-roll, Vox-style explainers, camera direction, captions, and creative-production workflows.

  • Seedance 2.0 prompt skill,使用该Skill生成Seedance 2.0 视频提示词

    Trust 82Quality 85media
  • Vox Director486 stars

    Turn one topic into a finished Vox-style paper-collage explainer/ad video — automated end to end on Atlas Clo…

    Trust 85Quality 95utility
  • OpenClaw/Codex Skill: Auto video editing for talk/vlog videos — speech recognition, sentence splitting, subti…

    Trust 81Quality 83media

Images, video, UI

Design and creative skills

Image, video, creative production, UI design, multimodal generation, and visual workflow skills.

  • Mlx Audio7.4K stars

    A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framewor…

    Trust 89Quality 100media-automation
  • A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality sin…

    Trust 86Quality 87media-automation
  • TTS Audio Suite1.0K stars

    A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion.…

    Trust 84Quality 93media-automation

Supply tracks

Build the registry by domain, not just by count.

Coding

Coding and developer agents

181

Code review, repo analysis, testing, CI, GitHub, DevOps, and developer workflow skills.

181 quality179 maintained

Research

Research and knowledge work

85

Deep research, source comparison, literature review, RAG, knowledge search, and reports.

85 quality75 maintained

Presentation

Presentation and deck workflows

1

PPTX generation, HTML slides, pitch decks, speaker notes, and presentation workflow skills.

1 quality1 maintained

Finance

Finance and quant workflows

14

Market data, SEC filings, portfolio analysis, quant research, backtesting, and risk workflows.

14 quality13 maintained

Marketing

Marketing and growth automation

10

SEO, content operations, lead generation, CRM, email automation, analytics, and growth workflows.

10 quality10 maintained

Design

Design and creative production

123

Design assets, images, video, audio, multimodal media, presentation, and creative production skills.

123 quality110 maintained

Data

Data, BI, and analytics

31

CSV, SQL, notebooks, dashboards, data pipelines, BI, ETL, and spreadsheet analysis.

31 quality31 maintained

Legal

Legal, policy, and compliance

13

Contract analysis, privacy, policy review, compliance checks, governance, and document risk review.

13 quality13 maintained

Education

Education and tutoring

5

Tutoring, course generation, quizzes, learning analytics, classrooms, and teaching workflows.

5 quality5 maintained

World Cup

Football and World Cup analytics

3

Football data, World Cup dashboards, xG, match prediction, scouting, and sports analytics.

3 quality3 maintained

High-intent entry points

Start from the task, not a keyword list.

These shortcuts use the same trust, supply, and relevance signals as the registry API, so humans and agents land on a useful shortlist faster.

Agent-readable index

Decision filters

Choose by scenario, quality, and trust signals.

Showing 1-16 of 454 ranked candidates matching "chinese-speech-recognition"

Best blend of quality, stars, freshness, and agent usage

1

ASRT SpeechRecognition

VERIFIEDEXCELLENT · 99TRUST · 86SAFE · REVIEWEDDESIGN

A Deep-Learning-Based Chinese Speech Recognition System 基于深度学习的中文语音识别系统

$ npx skills add nl8590687/ASRT_SpeechRecognition
8.4K stars66 quality86 trustReviewed4mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

pythonspeech
by nl8590687DetailsQuick view
2

Speech Recognition

VERIFIEDEXCELLENT · 100TRUST · 89SAFE · REVIEWEDDESIGN

Speech recognition module for Python, supporting several engines and APIs, online and offline.

$ npx skills add Uberi/speech_recognition
9.0K stars69 quality89 trustReviewed2mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

pythonspeech
by UberiDetailsQuick view
3

Expo Speech Recognition

STRONG · 76TRUST · 78SAFE · REVIEWEDDESIGN

Speech Recognition for React Native Expo projects

$ npx skills add jamsch/expo-speech-recognition
637 stars53 quality78 trustReviewed2mo since pushSafe to try
QualitySolid option that is likely worth shortlisting for production workflows.
TrustGood trust signals with a few areas worth checking before rollout.Review: Quality score needs review
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

typescriptspeech
by jamschDetailsQuick view
4

PaddleSpeech

VERIFIEDEXCELLENT · 100TRUST · 90SAFE · REVIEWEDDESIGN

Easy-to-use Speech Toolkit including Self-Supervised Learning model, SOTA/Streaming ASR with punctuation, Streaming TTS with text frontend, Speaker Verification System,…

$ npx skills add PaddlePaddle/PaddleSpeech
12.6K stars72 quality90 trustReviewed1mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

pythonspeech
by PaddlePaddleDetailsQuick view
5

Speechbrain

VERIFIEDEXCELLENT · 100TRUST · 88SAFE · REVIEWEDDESIGN

A PyTorch-based Speech Toolkit

$ npx skills add speechbrain/speechbrain
11.6K stars72 quality88 trustReviewed2mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

pythonspeech
by speechbrainDetailsQuick view
6

Speech Swift

STRONG · 78TRUST · 81SAFE · REVIEWEDDESIGN

AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML

$ npx skills add soniqo/speech-swift
894 stars54 quality81 trustReviewed2mo since pushSafe to try
QualitySolid option that is likely worth shortlisting for production workflows.
TrustGood trust signals with a few areas worth checking before rollout.Review: Quality score needs review
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

swiftspeech
by soniqoDetailsQuick view
7

Speech AI Forge

VERIFIEDEXCELLENT · 91TRUST · 85SAFE · REVIEWEDDESIGN

🍦 Speech-AI-Forge is a project developed around TTS generation model, implementing an API Server and a Gradio-based WebUI.

$ npx skills add lenML/Speech-AI-Forge
1.4K stars61 quality85 trustReviewed2mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustGood trust signals with a few areas worth checking before rollout.Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

pythonvoice
by lenMLDetailsQuick view
8

Fun ASR

VERIFIEDEXCELLENT · 94TRUST · 87SAFE · REVIEWEDDESIGN

End-to-end speech recognition large model: 31 languages, dialects, accents, lyrics, hotwords, timestamps, speaker diarization. Trained on tens of millions of hours.

$ npx skills add FunAudioLLM/Fun-ASR
1.3K stars64 quality87 trustReviewed1mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

cspeech
by FunAudioLLMDetailsQuick view
9

Whisper Finetune

VERIFIEDEXCELLENT · 96TRUST · 87SAFE · REVIEWEDDESIGN

Fine-tune the Whisper speech recognition model to support training without timestamp data, training with timestamp data, and training without speech data. Accelerate inf…

$ npx skills add yeyupiaoling/Whisper-Finetune
1.2K stars60 quality87 trustReviewed3mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

cspeech
by yeyupiaolingDetailsQuick view
10

Lhotse

VERIFIEDEXCELLENT · 93TRUST · 86SAFE · REVIEWEDDESIGN

Tools for handling multimodal data in machine learning projects.

$ npx skills add lhotse-speech/lhotse
1.1K stars63 quality86 trustReviewed1mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

pythonspeech
by lhotse-speechDetailsQuick view
11

Face Recognition

VERIFIEDSTRONG · 82TRUST · 79SAFE · EXPERIMENTALCODING

The world's simplest facial recognition api for Python and the command line

$ npx skills add ageitgey/face_recognition
56.5K stars61 quality79 trustExperimental2y since pushNeeds review
QualitySolid option that is likely worth shortlisting for production workflows.Check: Repository looks stale
TrustGood trust signals with a few areas worth checking before rollout.Review: Repository looks stale
Safety gateSparse or mixed signals. Useful for discovery, but not for autonomous installation.

Scenario GitHub automation

CLI + Codex · 4 targets

pythonmachine-learning
by ageitgeyDetailsQuick view
12

Whisper.Cpp

VERIFIEDEXCELLENT · 100TRUST · 88SAFE · REVIEWEDDESIGN

Port of OpenAI's Whisper model in C/C++

$ npx skills add ggml-org/whisper.cpp
51.0K stars76 quality88 trustReviewedOpenAI Agents1mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

OpenAI Agents + CLI · 4 targets

c++speech
by ggml-orgDetailsQuick view
13

WhisperX

VERIFIEDEXCELLENT · 100TRUST · 90SAFE · REVIEWEDDESIGN

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

$ npx skills add m-bain/whisperX
22.6K stars74 quality90 trustReviewed2mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

pythonspeech
by m-bainDetailsQuick view
14

Speech To Speech

VERIFIEDEXCELLENT · 99TRUST · 86SAFE · REVIEWEDDESIGN

Build local voice agents with open-source models

$ npx skills add huggingface/speech-to-speech
4.9K stars68 quality86 trustReviewed1mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

pythonmachine-learning
by huggingfaceDetailsQuick view
15

Chinese Copywriting Guidelines

VERIFIEDSTRONG · 78TRUST · 82SAFE · REVIEWEDRESEARCH

Chinese copywriting guidelines for better written communication/中文文案排版指北

$ npx skills add sparanoid/chinese-copywriting-guidelines
15.5K stars58 quality82 trustReviewed with permission notes1y since pushNeeds review
QualitySolid option that is likely worth shortlisting for production workflows.Check: Repository looks stale
TrustGood trust signals with a few areas worth checking before rollout.Review: Repository looks stale
Safety gateUsable candidate, but the agent should surface permission and audit notes before installation.

Scenario Document processing

CLI + Codex · 4 targets

markdown
by sparanoidDetailsQuick view
16

Vosk API

VERIFIEDEXCELLENT · 100TRUST · 88SAFE · REVIEWEDDESIGN

Offline speech recognition API for Android, iOS, Raspberry Pi and servers with Python, Java, C# and Node

$ npx skills add alphacep/vosk-api
14.9K stars72 quality88 trustReviewed2mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Multimodal media

CLI + Codex · 4 targets

jupyter-notebookspeech
by alphacepDetailsQuick view

Page 1

Showing the strongest 16 results to keep the registry fast for humans and agents. Refine by use case, platform, stars, or search query for a narrower shortlist.

Try the agent resolve API