Skill-Verzeichnis

Wiederverwendbare Skills für AI Agents entdecken.

Durchsuche reale GitHub-Skills nach Aufgabe und prüfe Stars, Trust, Audit, Kategorie und Installationspfad vor der Verwendung.

Jede Empfehlung bleibt mit ihrem Repository, Audit und Installationspfad nachvollziehbar.

Suchergebnisse: experiment

Englisches Verzeichnis
Aim82

Aim 💫 — An easy-to-use & supercharged open-source experiment tracker.

6.2K
Stars
82/100
Trust
Kategorie: ml-automationAudit

ClearML - Auto-Magical CI/CD to streamline your AI workload. Experiment Management, Data Management, Pipeline, Orchestration, Scheduling & Serving in one MLOps/LLMOps solution

6.7K
Stars
86/100
Trust
Kategorie: ml-automationAudit

An easy to use and powerful chaos engineering experiment toolkit.(阿里巴巴开源的一款简单易用、功能强大的混沌实验注入工具)

6.4K
Stars
86/100
Trust
Kategorie: devopsAudit

🧊 Open source LLM observability platform. One line of code to monitor, evaluate, and experiment. YC W23 🍓

6.0K
Stars
86/100
Trust
Kategorie: developmentAudit

The collaborative spreadsheet for AI. Chain cells into powerful pipelines, experiment with prompts and models, and evaluate LLM responses in real-time. Work together seamlessly to build and iterate on AI applications.

1.1K
Stars
84/100
Trust
Kategorie: rag-knowledgeAudit

A collection of Codex skills for IEEE-style academic writing, review, experiments, figures, LaTeX, citations, and paper reading.

114
Stars
73/100
Trust
Kategorie: researchAudit

How to use the Adaptyv Bio Foundry API and Python SDK for protein experiment design, submission, and results retrieval. Use this skill whenever the user mentions Adaptyv, Foundry API, protein binding assays, protein screening experiments, BLI/SPR assays, thermostability assays, or wants to submit protein sequences for experimental characterization. Also trigger when code imports `adaptyv`, `adaptyv_sdk`, or `FoundryClient`, or references `foundry-api-public.adaptyvbio.com`.

34K
Stars
78/100
Trust
Kategorie: design-creativeAudit

Autonomously improve a real artifact (code, training recipe, agent harness, data pipeline, prompt) against an objective and an evaluator, using Hypothesis Tree Refinement (HTR) from the Arbor paper. Use this whenever someone wants to iteratively optimize something over many experiments without overfitting — e.g. "get my model's eval score up", "improve this agent/harness", "tune this pipeline", "beat the baseline on this benchmark", "run a search over approaches and keep the best", "do an MLE-bench / Kaggle-style optimization", or any long-horizon "make this artifact better and don't just memorize the dev set" task. Trigger it even when the user doesn't say "Arbor" or "hypothesis tree" but describes repeated experiment-and-evaluate loops, branching exploration of competing ideas, or worries about a dev/test gap. Runs Claude itself as the coordinator with subagent executors in isolated git worktrees; for the standalone `arbor` CLI tool see references/arbor-upstream.md.

34K
Stars
77/100
Trust
Kategorie: researchAudit

When the user wants to plan, design, or implement an A/B test or experiment. Also use when the user mentions "A/B test," "split test," "experiment," "test this change," "variant copy," "multivariate test," "hypothesis," "conversion experiment," "statistical significance," or "test this." For tracking implementation, see analytics-tracking.

25K
Stars
77/100
Trust
Kategorie: design-creativeAudit

Determined is an open-source machine learning platform that simplifies distributed training, hyperparameter tuning, experiment tracking, and resource management. Works with PyTorch and TensorFlow.

3.2K
Stars
74/100
Trust
Kategorie: devopsAudit

The ultimate playground to learn, experiment with, and compare modern open-source AI agent frameworks — from basics to production-ready setups.

582
Stars
69/100
Trust
Kategorie: agent-frameworksAudit

Open Source ML Model Versioning, Metadata, and Experiment Management

1.7K
Stars
70/100
Trust
Kategorie: ml-automationAudit