Direktori skill

Temukan skill yang dapat digunakan kembali untuk AI agents.

Cari skill GitHub nyata berdasarkan tugas lalu periksa stars, trust, audit, kategori, dan jalur pemasangan sebelum digunakan.

Setiap rekomendasi tetap terhubung dengan repositori, audit, dan jalur pemasangannya.

Hasil pencarian: leaderboard

Direktori bahasa Inggris

SpeechIO Leaderboard: a large, robust, comprehensive, benchmarking platform for Automatic Speech Recognition.

545
Stars
63/100
Kepercayaan
Kategori: media-automationAudit

🛰️ A CLI tool for tracking token usage from OpenCode, Claude Code, 🦞OpenClaw (Clawdbot/Moltbot), Pi, Codex, Gemini, Cursor, AmpCode, Factory Droid, Kimi, and more! • 🏅Global Leaderboard + 2D/3D Contributions Graph

3.9K
Stars
74/100
Kepercayaan
Kategori: developmentAudit

35 production-grade agentic AI architectures (Reflexion, LATS, GraphRAG, MemGPT, Voyager, BrowserAgent, ...) — a Python library and runnable textbook with multi-provider LLM support and a 17-task benchmark leaderboard.

3.7K
Stars
85/100
Kepercayaan
Kategori: agent-frameworksAudit

Helps users discover and install agent skills when they ask questions like "how do I do X", "find a skill for X", "is there a skill that can...", or express interest in extending capabilities. This skill should be used when the user is looking for functionality that might exist as an installable skill.

51K
Stars
83/100
Kepercayaan
Kategori: researchAudit

Sharing both practical insights and theoretical knowledge about LLM evaluation that we gathered while managing the Open LLM Leaderboard and designing lighteval!

2.1K
Stars
73/100
Kepercayaan
Kategori: ml-automationAudit

A cross-platform AI agent skill that performs standardized health checkups with dual-axis scoring and public leaderboard integration.

95
Stars
67/100
Kepercayaan
Kategori: utilityAudit

Train and Infer Powerful Sentence Embeddings with AnglE | 🔥 SOTA on STS and MTEB Leaderboard

571
Stars
71/100
Kepercayaan
Kategori: rag-knowledgeAudit

🏆 The AI coding usage leaderboard — Claude Code, Codex, Gemini CLI & more. Real costs and tokens from ccusage data. Submit with: npx viberank-cli

102
Stars
68/100
Kepercayaan
Kategori: coding-agentsAudit

Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, leaderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具、基准/数据、演示、排行榜和大模型等组成的精选列表,主要面向基础大模型评测,旨在探求生成式AI的技术边界.

642
Stars
68/100
Kepercayaan
Kategori: rag-knowledgeAudit

A joint community effort to create one central leaderboard for LLMs.

306
Stars
62/100
Kepercayaan
Kategori: ml-automationAudit

READ this skill when designing or planning any game system architecture — including combat, skills, AI, UI, multiplayer, narrative, or scene systems. Contains paradigm selection guides (DDD / Data-Driven / Prototype), system-specific design references, and mixing strategies. Works as a domain knowledge plugin alongside workflow skills (OpenSpec, SpecKit) or plan mode of an agent.

57
Stars
57/100
Kepercayaan
Kategori: researchAudit

Cross-silo Federated Learning playground in Python. Discover 7 real-world federated datasets to test your new FL strategies and try to beat the leaderboard.

239
Stars
62/100
Kepercayaan
Kategori: geo-scienceAudit