Annuaire de skills

Découvrez des skills réutilisables pour les AI agents.

Recherchez de vrais skills GitHub par tâche et vérifiez Stars, confiance, audit, catégorie et chemin d’installation avant de les utiliser.

Chaque recommandation reste clairement reliée à son dépôt, son audit et son chemin d’installation.

Résultats de recherche: benchmark-datasets

Annuaire en anglais

🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23

29K
Stars
79/100
Confiance
Catégorie: developmentAudit

🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools

22K
Stars
87/100
Confiance
Catégorie: ml-automationAudit

Datasets, Transforms and Models specific to Computer Vision

18K
Stars
82/100
Confiance
Catégorie: ml-automationAudit

A data visualization and analytics component, especially well-suited for large and/or streaming datasets.

11K
Stars
87/100
Confiance
Catégorie: data-analysisAudit

Refine high-quality datasets and visual AI models

11K
Stars
82/100
Confiance
Catégorie: coding-agentsAudit

Full featured CSV parser with simple api and tested against large datasets.

4.3K
Stars
77/100
Confiance
Catégorie: data-analysisAudit

A Python library for anomaly detection across tabular, time series, graph, text, and image data. 60+ detectors, benchmark-backed ADEngine orchestration, and an agentic workflow for AI agents.

9.9K
Stars
86/100
Confiance
Catégorie: ml-automationAudit

Checks whether Kubernetes is deployed according to security best practices as defined in the CIS Kubernetes Benchmark

8.1K
Stars
86/100
Confiance
Catégorie: devopsAudit

A Python Package to Tackle the Curse of Imbalanced Datasets in Machine Learning

7.1K
Stars
86/100
Confiance
Catégorie: data-analysisAudit

A powerful tool for creating datasets for LLM fine-tuning 、RAG and Eval

14K
Stars
75/100
Confiance
Catégorie: dataAudit

非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat等商用模型, 以及step3.5-flash、kimi-k2.6、ernie4.5、MiniMax-M2.7、deepseek-v4、Qwen3.6、llama4、智谱GLM-5.1、MiMo-V2、LongCat、gemma4、mistral等开源大模型。不仅提供排行榜,也提供规模超200万的大模型缺陷库!方便广大社区研究分析、改进大模型。

6.2K
Stars
77/100
Confiance
Catégorie: agent-frameworksAudit

Argilla is a collaboration tool for AI engineers and domain experts to build high-quality datasets

5.0K
Stars
86/100
Confiance
Catégorie: ml-automationAudit