Openllmetry
traceloop
Open-source observability for your GenAI or LLM application, based on OpenTelemetry
OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 20
Candidates in this shortlist, not the full registry. GitHub stars belong to repositories, not individual skills.
Results: 20
traceloop
Open-source observability for your GenAI or LLM application, based on OpenTelemetry
langfuse
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK,…
comet-ml
Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
0xPlaygrounds
⚙️🦀 Build modular and scalable LLM Applications in Rust
bentoml
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
katanemo
Plano is an AI-native proxy and data plane for agentic apps — with built-in orchestration, safety, observability, and smart LLM routing so you stay focused on your agent…
Helicone
🧊 Open source LLM observability platform. One line of code to monitor, evaluate, and experiment. YC W23 🍓
Agenta-AI
The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.
Giskard-AI
🐢 Open-Source Evaluation & Testing library for LLM Agents
truera
Evaluation and Tracking for LLM Experiments and AI Agents
JuliusBrussee
🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman
Leonxlnx
Taste-Skill - gives your AI good taste. stops the AI from generating boring, generic slop
conorbronsdon
Skill that audits and rewrites content to remove AI writing patterns. Use it with your favorite agents including Claude Code, OpenClaw, and Hermes.