Skill 审计报告
Chinese Llm Benchmark 审计报告.
非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat等商用模型, 以及step3.5-flash、kimi-k2.6、ernie4.5、MiniMax-M2.7、deepseek-v4、Qwen3.6、llama4、智谱GLM-5.1、MiMo-V2、LongCat、gemma4、mistral等开源大模型。不仅提供排行榜,也提供规模超200万的大模型缺陷库!方便广大社区研究分析、改进大模型。
OpenAgentSkill 信任评分
OpenAgentSkill 信任评分
Trust Score 帮助 Agent 在安装前判断一个 Skill 是否足以进入候选清单。
GitHub 采用度
通过94
6.2K 个 GitHub Stars
Star/Fork 活跃度
通过88
6.2K 个 Star,250 个 Fork; 当前元数据中没有议题活跃度信息
近期维护
通过88
距上次推送 3 个月
许可证清晰度
警告42
未知
README/SKILL.md 完整度
通过90
元数据包含足够的用法与工作流上下文
依赖与运行时风险
通过90
公开元数据中未发现主要依赖风险提示
安装可用性
通过92
npx skills add jeinlee1991/chinese-llm-benchmark
安装命令安全性
通过92
标准软件包或运行时安装路径
权限范围
通过86
文件系统或文档访问
仓库证据
通过86
https://github.com/jeinlee1991/chinese-llm-benchmark
审查状态
通过88
可用 AI 审查数据
Agent 验证结果
信息54
暂未有 Agent 结果数据
检查项
安装与采用审查
安装路径
92
npx skills add jeinlee1991/chinese-llm-benchmark
仓库
88
https://github.com/jeinlee1991/chinese-llm-benchmark
许可证
45
未知
维护
88
距上次推送 3 个月
AI 审查
88
Approved with no listed issues
README/SKILL.md 完整度
90
Usable description available
依赖风险
90
公开元数据中未发现主要依赖风险提示
安装命令安全性
92
标准软件包或运行时安装路径
权限范围
86
文件系统或文档访问
Star/Fork 活跃度
88
6.2K 个 Star,250 个 Fork; 当前元数据中没有议题活跃度信息
采用度
88
6.2K 个 GitHub Stars
Financial decision safety
58
Research-only use: do not treat output as financial advice or execute a position without human approval.
警告
- 许可证不清晰
- Financial research output is not financial advice; require human review before any live investment decision
- Financial research output is not financial advice; require human review before any live investment decision.
- License clarity: Unknown
方法
本报告综合公开元数据、AI 审查输出、仓库活跃度、安装就绪度、OpenAgentSkill 事件、质量评分、信任检查和 Agent 安全门槛;它不是完整的源代码安全审计。
对比相近选项