跳到主要内容
#90
该品类排名 #90 / 共 93 个· Ranked #90 of 93 in this category

agent-evaluation

by NeoLabHQ·2 个月前

Evaluate and improve Claude Code commands, skills, and agents. Use when testing prompt effectiveness, validating context

测试与调试Claude CodeCodex自动筛查 · automated开源 · open source

安装前须知Before you install

未声明特殊权限需求No special access needs declared
编辑结论· Editor's verdict

Auto-published by GitHub discovery based on star threshold.

— 编辑团队 · Editorial team

通过 Skills CLI 安装

使用 npx skills add 将该 skill 安装到选中的 Agent。Phase 0 命令均为按规则生成,尚未验证。

Codex
npx skills add https://github.com/NeoLabHQ/context-engineering-kit/tree/master/plugins/customaize-agent/skills/agent-evaluation -g -a codex -y

去掉 -g 可改为项目级安装

适合什么场景Best for

  • High-star GitHub skill discovery candidate.

不适合什么场景Not for

  • Sensitive or production workflows without local review.

vs 其他选择vs alternatives

完整对比表Full compare table →

维度对比side-by-side compare

和同类的关键维度差异
当前 · this skillagent-evaluationQA LoopGStack InvestigateGStack QA
评分 · rating4.94.84.8
星标 · stars1.0k15k123k123k
风险 · risk自动筛查 · automated中风险 · med risk低风险 · low risk中风险 · med risk
最适合 · best forHigh-star GitHub skill discovery candidate.浏览器冒烟测试结构化调试浏览器回归循环
不适合 · not forSensitive or production workflows without local review.纯单元测试快速文案改动只读审计环境

审计备注Audit notes

未独立审计 · not audited
源码Source公开 GitHub · open on GitHub
作者Author社区贡献 · community!
网络访问Network未独立审计 · not audited!
文件写入Filesystem未独立审计 · not audited!
依赖Dependencies未独立审计 · not audited!
遥测Telemetry无 · none
Skill Market
按品类找最好用的 AI 技能·Find the best AI skills for the job
v0.4 · 收录 1306 个 skill · 上次评测 2026-06-10