AI Agent Skill Repository

AI Agent Skills Directory

Browse reusable skills for Codex, Claude Code, Cursor, finance, research, web scraping, PPT, football analytics, data, marketing, design, and more.

Browse by scenario

Real skills, grouped by the work your agent needs to finish.

Each directory entry links to real skill pages with GitHub adoption, trust score, install handoff, risk notes, and agent-readable metadata. Use these as starting points when you want a shortlist before asking an agent to install anything.

Design to deployment

Frontend and UI skills

Design direction, Figma implementation, UI review, React performance, browser QA, and safe preview deployments.

  • Frontend Design166K stars

    Guidance for distinctive, intentional UI design, typography, visual direction, and non-template-like product…

    Trust 88Quality 100design-creative
  • A set of skills including RAWeb, RAM, and RAPDF coverage for AI (targeting Claude Code)

    Trust 74Quality 66coding-agents
  • Dokploy36K stars

    Open Source Alternative to Vercel, Netlify and Heroku.

    Trust 85Quality 100devops

Codex, Claude Code, Cursor

Coding agent skills

Code review, repo inspection, testing, planning, shipping, and engineering workflows for coding agents.

  • Code Review169K stars

    Review a diff separately for code standards and spec compliance.

    Trust 91Quality 100coding-agents
  • Agent Skills81K stars

    Production-grade engineering skills for AI coding agents.

    Trust 86Quality 100agent-skills
  • Implement176K stars

    Execute an approved engineering ticket with tests, review, and a branch commit.

    Trust 88Quality 100coding-agents

Documents and knowledge

Research and RAG skills

Research, recent-events briefings, PDF parsing, markdown conversion, RAG ingestion, and knowledge workflows.

  • Markitdown156K stars

    Python tool for converting files and office documents to Markdown.

    Trust 87Quality 100document-processing
  • AutoRAG4.8K stars

    AutoRAG: An Open-Source Framework for Retrieval-Augmented Generation (RAG) Evaluation & Optimization with Aut…

    Trust 87Quality 99data
  • FinSight AI1.2K stars

    AI equity research agent with resilient workflows, Redis Lua single-flight, pgvector RAG, versioned reports,…

    Trust 87Quality 93rag-knowledge

Markets and quant

Finance and trading skills

Stock analysis, market research, quant backtesting, financial data, and investment research skills.

  • Stocks375 stars

    Programs for stock prediction and evaluation

    Trust 77Quality 74finance
  • 入门资料整理:1.多因子股票量化框架开源教程 2.学界和业界的经典资料收录 3.AI + 金融的相关工作,包括LLM, Agent, benchmark(evaluation), etc.

    Trust 87Quality 94finance
  • Alphasift185 stars

    AI-native stock screening engine with full-market discovery, LLM ranking, risk-aware scoring, and auditable e…

    Trust 78Quality 71finance

Crawlers and extraction

Web scraping skills

Crawling, scraping, extraction, browser automation, structured data capture, and website-to-markdown workflows.

  • Firecrawl139K stars

    The API to search, scrape, and interact with the web at scale. 🔥

    Trust 87Quality 100agent-frameworks
  • Crawl4AI73K stars

    Web crawling built for AI

    Trust 90Quality 100web-automation
  • Bisheng11K stars

    BISHENG is an open LLM devops platform for next generation Enterprise AI applications. Powerful and comprehen…

    Trust 90Quality 100document-processing

Slides and decks

PPT and presentation skills

Presentation generation, editable PPTX, slide decks, speaker notes, and visual storytelling workflows.

  • Reverse-lookup glossary that turns a vague description of a web animation or motion effect into its exact ter…

    Trust 91Quality 100design-creative
  • Evidence-driven Agent Skills for design award research, evaluation, award matching, entry writing, and submis…

    Trust 86Quality 87design-creative

Prompts, B-roll, explainers

Video creation skills

Video-generation prompts, B-roll, Vox-style explainers, camera direction, captions, and creative-production workflows.

  • VideoGen Eval266 stars

    VideoGen-Eval: Agent-based System for Video Generation Evaluation

    Trust 71Quality 57media-automation
  • Video ChatGPT1.5K stars

    [ACL 2024 🔥] Video-ChatGPT is a video conversation model capable of generating meaningful conversation about…

    Trust 84Quality 85support-automation

Images, video, UI

Design and creative skills

Image, video, creative production, UI design, multimodal generation, and visual workflow skills.

  • Wrapper for 10+ VPR models. Use any SOTA VPR model just by changing one parameter

    Trust 77Quality 68robotics-iot
  • IntelliNode274 stars

    Access the latest AI models like ChatGPT, LLaMA, Deepseek, Diffusion, Hugging face, and beyond through a unif…

    Trust 75Quality 58media-automation
  • ImagenHub180 stars

    A one-stop library to standardize the inference and evaluation of all the conditional image generation models…

    Trust 75Quality 56media-automation

Supply tracks

Build the registry by domain, not just by count.

Coding

Coding and developer agents

207

Code review, repo analysis, testing, CI, GitHub, DevOps, and developer workflow skills.

205 quality199 maintained

Research

Research and knowledge work

85

Deep research, source comparison, literature review, RAG, knowledge search, and reports.

83 quality78 maintained

Presentation

Presentation and deck workflows

1

PPTX generation, HTML slides, pitch decks, speaker notes, and presentation workflow skills.

1 quality1 maintained

Finance

Finance and quant workflows

17

Market data, SEC filings, portfolio analysis, quant research, backtesting, and risk workflows.

17 quality17 maintained

Marketing

Marketing and growth automation

10

SEO, content operations, lead generation, CRM, email automation, analytics, and growth workflows.

10 quality10 maintained

Design

Design and creative production

58

Design assets, images, video, audio, multimodal media, presentation, and creative production skills.

57 quality51 maintained

Data

Data, BI, and analytics

33

CSV, SQL, notebooks, dashboards, data pipelines, BI, ETL, and spreadsheet analysis.

32 quality32 maintained

Legal

Legal, policy, and compliance

12

Contract analysis, privacy, policy review, compliance checks, governance, and document risk review.

11 quality11 maintained

Education

Education and tutoring

5

Tutoring, course generation, quizzes, learning analytics, classrooms, and teaching workflows.

5 quality5 maintained

World Cup

Football and World Cup analytics

3

Football data, World Cup dashboards, xG, match prediction, scouting, and sports analytics.

3 quality2 maintained

High-intent entry points

Start from the task, not a keyword list.

These shortcuts use the same trust, supply, and relevance signals as the registry API, so humans and agents land on a useful shortlist faster.

Agent-readable index

Decision filters

Choose by scenario, quality, and trust signals.

Showing 1-16 of 417 ranked candidates matching "ai-evaluation"

Best blend of quality, stars, freshness, and agent usage

1

Evaluation Guidebook

VERIFIEDSTRONG · 76TRUST · 81SAFE · REVIEWEDDESIGN

Sharing both practical insights and theoretical knowledge about LLM evaluation that we gathered while managing the Open LLM Leaderboard and designing lighteval!

$ npx skills add huggingface/evaluation-guidebook
2.1K stars54 quality81 trustReviewed with permission notes8mo since pushNeeds review
QualitySolid option that is likely worth shortlisting for production workflows.
TrustGood trust signals with a few areas worth checking before rollout.Review: License is unclear
Safety gateUsable candidate, but the agent should surface permission and audit notes before installation.

Scenario Design and creative

CLI + Codex · 4 targets

jupyter-notebookmachine-learning
by huggingfaceDetailsQuick view
2

Cloud Map Evaluation

PROMISING · 63TRUST · 73SAFE · REVIEWEDCODING

[RAL' 25 & IROS‘ 25] MapEval: Towards Unified, Robust and Efficient SLAM Map Evaluation Framework.

$ npx skills add JokerJohn/Cloud_Map_Evaluation
473 stars45 quality73 trustReviewed with permission notes4mo since pushNeeds review
QualityUseful candidate, but compare it with alternatives before adopting.
TrustGood trust signals with a few areas worth checking before rollout.Review: License is unclear
Safety gateUsable candidate, but the agent should surface permission and audit notes before installation.

Scenario GitHub automation

CLI + Codex · 4 targets

c++robotics
by JokerJohnDetailsQuick view
3

Place Recognition Evaluation

NEEDS REVIEW · 36TRUST · 67SAFE · BLOCKEDAUTOMATION

Benchmarking and evaluation framework for place recognition methods, featuring SuperPoint+SuperGlue, LoGG3D-Net, Scan Context, DBoW2, MixVPR, STD

$ npx skills add 4ku/Place-recognition-evaluation
126 stars33 quality67 trustBlocked for auto-install2y since pushRisky
QualityInspect the repository carefully before adding it to an agent workflow.Check: Repository looks stale
TrustPotentially useful, but at least one trust signal needs human inspection.Review: License is unclear
Safety gateThis skill should not be selected by an agent without explicit human security review.

Scenario Workflow automation

CLI + Codex · 4 targets

c++robotics
by 4kuDetailsQuick view
4

RagaAI Catalyst

VERIFIEDEXCELLENT · 100TRUST · 89SAFE · REVIEWEDCODING

Python SDK for Agent AI Observability, Monitoring and Evaluation Framework. Includes features like agent, llm and tools tracing, debugging multi-agentic system, self-hos…

$ npx skills add raga-ai-hub/RagaAI-Catalyst
16.1K stars66 quality89 trustReviewed6mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Coding agents

CLI + Codex · 4 targets

pythonllmops
by raga-ai-hubDetailsQuick view
5

VPR Methods Evaluation

PROMISING · 68TRUST · 77SAFE · REVIEWEDDESIGN

Wrapper for 10+ VPR models. Use any SOTA VPR model just by changing one parameter

$ npx skills add gmberton/VPR-methods-evaluation
197 stars47 quality77 trustReviewed with permission notes3mo since pushNeeds review
QualityUseful candidate, but compare it with alternatives before adopting.
TrustGood trust signals with a few areas worth checking before rollout.Review: Quality score needs review
Safety gateUsable candidate, but the agent should surface permission and audit notes before installation.

Scenario Multimodal media

CLI + Codex · 4 targets

pythoncomputer-vision
by gmbertonDetailsQuick view
6

Langfuse

VERIFIEDEXCELLENT · 100TRUST · 87SAFE · REVIEWEDCODING

🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK,…

$ npx skills add langfuse/langfuse
29.4K stars74 quality87 trustReviewedOpenAI Agents + LangChain1mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: License is unclear
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Coding agents

OpenAI Agents + LangChain · 4 targets

typescriptllmops
by langfuseDetailsQuick view
7

Mlflow

VERIFIEDEXCELLENT · 100TRUST · 90SAFE · REVIEWEDCODING

The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality A…

$ npx skills add mlflow/mlflow
26.7K stars74 quality90 trustReviewedLangChain1mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Coding agents

LangChain + CLI · 4 targets

pythonllmops
by mlflowDetailsQuick view
8

Opik

VERIFIEDEXCELLENT · 100TRUST · 90SAFE · REVIEWEDCODING

Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.

$ npx skills add comet-ml/opik
19.8K stars73 quality90 trustReviewedLangChain1mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Coding agents

LangChain + CLI · 4 targets

pythonllmops
by comet-mlDetailsQuick view
9

Crawl4AI

VERIFIEDEXCELLENT · 100TRUST · 90SAFE · REVIEWEDRESEARCH

Web crawling built for AI

$ npx skills add unclecode/crawl4ai
4 agent calls100% success73.1K stars80 quality90 trustReviewedClaude Code + OpenAI Agents18d since pushSafe to try31.0K installs
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Research agents

Claude Code + OpenAI Agents · 4 targets

claudegpt-4langchaincrewaiopenclaw
by unclecodeDetailsQuick view
10

Giskard Oss

VERIFIEDEXCELLENT · 100TRUST · 88SAFE · REVIEWEDCODING

🐢 Open-Source Evaluation & Testing library for LLM Agents

$ npx skills add Giskard-AI/giskard-oss
5.7K stars68 quality88 trustReviewed16d since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Coding agents

CLI + Codex · 4 targets

pythonllmops
by Giskard-AIDetailsQuick view
11

Coze Loop

VERIFIEDEXCELLENT · 100TRUST · 90SAFE · REVIEWEDCODING

Next-generation AI Agent Optimization Platform: Cozeloop addresses challenges in AI agent development by providing full-lifecycle management capabilities from developmen…

$ npx skills add coze-dev/coze-loop
5.6K stars68 quality90 trustReviewed20d since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Coding agents

CLI + Codex · 4 targets

gollmops
by coze-devDetailsQuick view
12

AutoRAG

VERIFIEDEXCELLENT · 99TRUST · 87SAFE · REVIEWEDRESEARCH

AutoRAG: An Open-Source Framework for Retrieval-Augmented Generation (RAG) Evaluation & Optimization with AutoML-Style Automation

$ npx skills add Marker-Inc-Korea/AutoRAG
4.8K stars67 quality87 trustReviewed2mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario RAG and knowledge

CLI + Codex · 4 targets

pythonrag
by Marker-Inc-KoreaDetailsQuick view
13

VLMEvalKit

VERIFIEDEXCELLENT · 99TRUST · 88SAFE · REVIEWEDRESEARCH

Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks

$ npx skills add open-compass/VLMEvalKit
4.2K stars67 quality88 trustReviewedClaude Code + OpenAI Agents2mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Research agents

Claude Code + OpenAI Agents · 4 targets

pythoncomputer-vision
by open-compassDetailsQuick view
14

Agenta

VERIFIEDEXCELLENT · 94TRUST · 84SAFE · REVIEWEDCODING

The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.

$ npx skills add Agenta-AI/agenta
4.2K stars67 quality84 trustReviewed2mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustGood trust signals with a few areas worth checking before rollout.Review: License is unclear
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Coding agents

CLI + Codex · 4 targets

typescriptllmops
by Agenta-AIDetailsQuick view
15

Trulens

VERIFIEDEXCELLENT · 98TRUST · 86SAFE · REVIEWEDCODING

Evaluation and Tracking for LLM Experiments and AI Agents

$ npx skills add truera/trulens
3.4K stars66 quality86 trustReviewed2mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Coding agents

CLI + Codex · 4 targets

pythonllmops
by trueraDetailsQuick view
16

Langwatch

VERIFIEDEXCELLENT · 98TRUST · 86SAFE · REVIEWEDCODING

The platform for LLM evaluations and AI agent testing

$ npx skills add langwatch/langwatch
3.3K stars66 quality86 trustReviewedOpenAI Agents2mo since pushSafe to try
QualityHigh-confidence pick with strong adoption and healthy maintenance signals.
TrustStrong OpenAgentSkill Trust Score across adoption, recent maintenance, license clarity, documentation, dependency/…Review: Documentation summary is thin
Safety gateGood audit and safety signals with no high-risk permission hints in public metadata.

Scenario Coding agents

OpenAI Agents + CLI · 4 targets

typescriptllmops
by langwatchDetailsQuick view

Page 1

Showing the strongest 16 results to keep the registry fast for humans and agents. Refine by use case, platform, stars, or search query for a narrower shortlist.

Try the agent resolve API