Directorio de skills

Descubre skills reutilizables para AI agents.

Busca skills reales de GitHub por tarea y revisa stars, confianza, auditoría, categoría y ruta de instalación antes de utilizarlos.

Cada recomendación conserva un vínculo claro con su repositorio, auditoría y ruta de instalación.

Resultados de búsqueda: benchmark-datasets

Directorio en inglés

🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK, LiteLLM, and more. 🍊YC W23

29K
Stars
79/100
Confianza
Categoría: developmentAuditoría

🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools

22K
Stars
87/100
Confianza
Categoría: ml-automationAuditoría

Datasets, Transforms and Models specific to Computer Vision

18K
Stars
82/100
Confianza
Categoría: ml-automationAuditoría

A data visualization and analytics component, especially well-suited for large and/or streaming datasets.

11K
Stars
87/100
Confianza
Categoría: data-analysisAuditoría

Refine high-quality datasets and visual AI models

11K
Stars
82/100
Confianza
Categoría: coding-agentsAuditoría

Full featured CSV parser with simple api and tested against large datasets.

4.3K
Stars
77/100
Confianza
Categoría: data-analysisAuditoría

A Python library for anomaly detection across tabular, time series, graph, text, and image data. 60+ detectors, benchmark-backed ADEngine orchestration, and an agentic workflow for AI agents.

9.9K
Stars
86/100
Confianza
Categoría: ml-automationAuditoría

Checks whether Kubernetes is deployed according to security best practices as defined in the CIS Kubernetes Benchmark

8.1K
Stars
86/100
Confianza
Categoría: devopsAuditoría

A Python Package to Tackle the Curse of Imbalanced Datasets in Machine Learning

7.1K
Stars
86/100
Confianza
Categoría: data-analysisAuditoría

A powerful tool for creating datasets for LLM fine-tuning 、RAG and Eval

14K
Stars
75/100
Confianza
Categoría: dataAuditoría

非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat等商用模型, 以及step3.5-flash、kimi-k2.6、ernie4.5、MiniMax-M2.7、deepseek-v4、Qwen3.6、llama4、智谱GLM-5.1、MiMo-V2、LongCat、gemma4、mistral等开源大模型。不仅提供排行榜,也提供规模超200万的大模型缺陷库!方便广大社区研究分析、改进大模型。

6.2K
Stars
77/100
Confianza
Categoría: agent-frameworksAuditoría

Argilla is a collaboration tool for AI engineers and domain experts to build high-quality datasets

5.0K
Stars
86/100
Confianza
Categoría: ml-automationAuditoría