No visual example yet
Explore the skillChinese Llm Benchmark
jeinlee1991
非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat…
OPENAGENTSKILL / DIRECTORY
次のタスクに合うスキルを。Codex、Claude Code、Cursor などのツールを探せます。
110 Skills
検索結果: 110
No visual example yet
Explore the skilljeinlee1991
非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat…
No visual example yet
Explore the skillShubhamsaboo
100+ AI Agent & RAG apps you can actually run — clone, customize, ship.
No visual example yet
Explore the skillMintplex-Labs
Stop renting your intelligence. Own it with AnythingLLM. Everything you need for a powerful local-first agent experience
No visual example yet
Explore the skillpathwaycom
Ready-to-run cloud templates for RAG, AI pipelines, and enterprise search with live data. 🐳Docker-friendly.⚡Always in sync with Sharepoint, Google Drive, S3, Kafka, Pos…
No visual example yet
Explore the skilldatawhalechina
No visual example yet
Explore the skillSamurAIGPT
A personal knowledge base that builds and maintains itself. Drop in sources — Claude (or Codex/Gemini) reads them, extracts knowledge, and maintains a persistent interli…
No visual example yet
Explore the skillmodelscope
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
No visual example yet
Explore the skilllangfuse
🪢 Open source AI engineering platform: LLM evals, observability, metrics, prompt management, playground, datasets. Integrates with OpenTelemetry, LangChain, OpenAI SDK,…
No visual example yet
Explore the skillcomet-ml
Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
No visual example yet
Explore the skillAgenta-AI
The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.
No visual example yet
Explore the skillGiskard-AI
🐢 Open-Source Evaluation & Testing library for LLM Agents
No visual example yet
Explore the skilltruera
Evaluation and Tracking for LLM Experiments and AI Agents
No visual example yet
Explore the skilltraceloop
Open-source observability for your GenAI or LLM application, based on OpenTelemetry
No visual example yet
Explore the skillkaushikb11
No visual example yet
Explore the skilldatawhalechina
本项目是一个面向小白开发者的大模型应用开发教程,在线阅读地址:https://datawhalechina.github.io/llm-universe/
No visual example yet
Explore the skillneo4j-labs