OPENAGENTSKILL / DIRECTORY
Find a skill for your next task. Explore tools for Codex, Claude Code, Cursor and more.
1–16 / 178
Results: 178
seleniumbase
SeleniumBase is a framework for UI Testing, Web Scraping, and Stealth. Passes every bot-detection test with CDP Mode, and extends Playwright.
Giskard-AI
🐢 Open-Source Evaluation & Testing library for LLM Agents
traceloop
Open-source observability for your GenAI or LLM application, based on OpenTelemetry
mattpocock
Pressure-test a plan against the codebase, domain glossary, and architectural decisions.
alchaincyf
达尔文.skill —— 一个让你的Skill无限进化的系统:评估→改进→测试→保留或回滚 | Autoresearch-inspired autonomous skill optimization for Claude Code. Evaluate, improve, test, keep or revert.
anthropics
Use Playwright to interact with and test local web applications, capture screenshots, debug UI behavior, and inspect browser logs.
rtk-ai
CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies
vercel-labs
React and Next.js performance guidance for writing, reviewing, and refactoring production UI code.
langfuse
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK,…
comet-ml
Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
Owner-curated external sources. Not filtered by the scores or compatibility controls above; excluded from GitHub rankings and automatic installation.
No external entries match this query.