OPENAGENTSKILL / DIRECTORY

AI Agent Skills

Finde den passenden Skill für deine nächste Aufgabe mit Codex, Claude Code, Cursor und mehr.

Ergebnisse · “llm-evaluation”

109 Skills

Ergebnisse: 109

KI & Wissen

No visual example yet

Explore the skill

非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat…

Preis unbestätigtKI & WissenClaude CodeOpenAI AgentsVor Nutzung prüfen
6163GitHub
Skill ansehen
KI & Wissen

No visual example yet

Explore the skill

Langfuse

langfuse

🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK,…

Preis unbestätigtKI & WissenOpenAI AgentsLangChainVor Nutzung prüfen
29.437GitHub
Skill ansehen
Entwicklung

No visual example yet

Explore the skill

Opik

comet-ml

Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.

Preis unbestätigtEntwicklungLangChainVor Nutzung prüfen
19.788GitHub
Skill ansehen
KI & Wissen

No visual example yet

Explore the skill

Agenta

Agenta-AI

The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.

Preis unbestätigtKI & WissenVor Nutzung prüfen
4450GitHub
Skill ansehen
KI & Wissen

No visual example yet

Explore the skill

Openllmetry

traceloop

Open-source observability for your GenAI or LLM application, based on OpenTelemetry

Preis unbestätigtKI & WissenVor Nutzung prüfen
7246GitHub
Skill ansehen
KI & Wissen

No visual example yet

Explore the skill

AutoRAG

Marker-Inc-Korea

AutoRAG: An Open-Source Framework for Retrieval-Augmented Generation (RAG) Evaluation & Optimization with AutoML-Style Automation

Preis unbestätigtKI & WissenVor Nutzung prüfen
4829GitHub
Skill ansehen

Anleitungen & Vergleiche

Für Entwickler